Skip to main content

Gemini 3.1 Pro Preview

GoogleGeminiProprietary

Description

Gemini 3.1 Pro is the latest model in the Gemini 3 series. It excels at complex tasks requiring broad world knowledge and advanced reasoning across modalities. Gemini 3.1 Pro uses dynamic thinking by default to reason through prompts, and features a 1 million-token input context window with 64k output tokens.

Release Date
2026-02-19
Parameters
—
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

33
general
67
coding
94
reasoning
72
science
80
agents
80
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability19
60.0
LS
Code Ranking90
86.0
AA
General Ranking69
71.0
AA
Math Reasoning27
91.0
LB
Multimodal Ranking32
60.0
LS
Reasoning34
84.0
LB
Science26
86.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench99.3%SR
Finance Agent v243.0%

Code

FrontierSWE40.0%
DeepSWE 1.112.0%

Language

MMMLU92.6%SR

Legal

Legal Agent Benchmark0.0%

Long Context

MRCR v2 (8-needle)26.3%SR

Math

LiveBench79.9%

Reasoning

LiveCodeBench Pro2887.00 / 3000SR
GPQANYU + Cohere + Anthropic (2023)94.3%SR
BrowseCompOpenAI (2025)85.9%SR
SWE-Bench Verified80.6%SR
ARC-AGI v277.1%SR
MCP Atlas69.2%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)68.5%SR
SciCode59.0%SR
SWE-Bench ProPrinceton NLP (2024)54.2%SR
Humanity's Last Exam51.4%SR
APEX-Agents33.5%SR

Vision

MMMU-Pro80.5%SR

AA Evaluation Indices

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
95.6
Gpqa(NYU + Cohere + Anthropic (2023))
94.1
Lcr(Artificial Analysis)
82.0
Ifbench(Google Research (2023))
77.1
Terminalbench V2 1
73.8
Coding Index(Artificial Analysis)
68.8
Scicode(UIUC + Argonne National Lab (2024))
58.7
Terminalbench Hard(Stanford × Laude Institute (2026))
53.8
Hle(Center for AI Safety + Scale AI (2025))
47.0
Intelligence Index(Artificial Analysis)
29.7
Tau Banking
21.4
Terminalbench V4 0
4.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Code
100
Reasoning
100
General
100
Language
90
Search
90
Multimodal
80
Physics
80
Spatial Reasoning
80
Frontend Development
80
Biology
80
Chemistry
80
Tool Calling
80
Math
70
Vision
70
Agents
50
Finance
40
Long Context
30
Healthcare
20
Legal
0

Pricing

Input Price$2 / 1M tokens
Output Price$12 / 1M tokens
Blended Price (3:1)$4.5 / 1M tokens
Cache Read Price$0.2 / 1M tokens

Speed

Tokens/sec129.6
Time to First Token26.24s
Time to Answer26.24s

Provider Price Ranking

Provider Price Ranking

27 providers

Cheapest: DeepInfraMost Expensive: Venice AI
ProviderInputOutput
1DeepInfraCheapest
$0
$0.00001
2Google
$0
$0.00002
3Kilo Gateway
$1
$6
4302.AI
$2
$12
5NanoGPT
$2
$12
6Abacus
$2
$12
7Perplexity Agent
$2
$12
8OpenRouter
$2
$12
9ZenMux
$2
$12
10Vivgrid
$2
$12
11FrogBot
$2
$12
12AIHubMix
$2
$12
13Requesty
$2
$12
14Vercel AI Gateway
$2
$12
15DevPass (LLM Gateway)
$2
$12
16Vertex
$2
$12
17FastRouter
$2
$12
18Auriko
$2
$12
19OrcaRouter
$2
$12
20Merge Gateway
$2
$12
21DaoXE
$2
$12
22Ofox
$2
$12
23Impossibl
$2
$12
24Eden AI
$2
$12
25Opper
$2
$12
26Tempr Gateway
$2
$12
27Venice AI
$2.5
$15

Compare pricing across different API providers for this model.

External Sources