Gemini 3.1 Pro Preview
GoogleGeminiProprietary
Description
Gemini 3.1 Pro is the latest model in the Gemini 3 series. It excels at complex tasks requiring broad world knowledge and advanced reasoning across modalities. Gemini 3.1 Pro uses dynamic thinking by default to reason through prompts, and features a 1 million-token input context window with 64k output tokens.
Release Date
2026-02-19
Parameters
—
Context Length
1.0M
Modalities
audio, image, pdf, text, video
Capability Radar
33
general
67
coding
94
reasoning
72
science
80
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 19 | 60.0 | LS |
| Code Ranking | 90 | 86.0 | AA |
| General Ranking | 69 | 71.0 | AA |
| Math Reasoning | 27 | 91.0 | LB |
| Multimodal Ranking | 32 | 60.0 | LS |
| Reasoning | 34 | 84.0 | LB |
| Science | 26 | 86.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
99.3%SR
Finance Agent v2
43.0%
Code
FrontierSWE
40.0%
DeepSWE 1.1
12.0%
Language
MMMLU
92.6%SR
Legal
Legal Agent Benchmark
0.0%
Long Context
MRCR v2 (8-needle)
26.3%SR
Math
LiveBench
79.9%
Reasoning
LiveCodeBench Pro
2887.00 / 3000SR
GPQANYU + Cohere + Anthropic (2023)
94.3%SR
BrowseCompOpenAI (2025)
85.9%SR
SWE-Bench Verified
80.6%SR
ARC-AGI v2
77.1%SR
MCP Atlas
69.2%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)
68.5%SR
SciCode
59.0%SR
SWE-Bench ProPrinceton NLP (2024)
54.2%SR
Humanity's Last Exam
51.4%SR
APEX-Agents
33.5%SR
Vision
MMMU-Pro
80.5%SR
AA Evaluation Indices
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))95.6
Gpqa(NYU + Cohere + Anthropic (2023))94.1
Lcr(Artificial Analysis)82.0
Ifbench(Google Research (2023))77.1
Terminalbench V2 173.8
Coding Index(Artificial Analysis)68.8
Scicode(UIUC + Argonne National Lab (2024))58.7
Terminalbench Hard(Stanford × Laude Institute (2026))53.8
Hle(Center for AI Safety + Scale AI (2025))47.0
Intelligence Index(Artificial Analysis)29.7
Tau Banking21.4
Terminalbench V4 04.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Code100
Reasoning100
General100
Language90
Search90
Multimodal80
Physics80
Spatial Reasoning80
Frontend Development80
Biology80
Chemistry80
Tool Calling80
Math70
Vision70
Agents50
Finance40
Long Context30
Healthcare20
Legal0
Pricing
Input Price$2 / 1M tokens
Output Price$12 / 1M tokens
Blended Price (3:1)$4.5 / 1M tokens
Cache Read Price$0.2 / 1M tokens
Speed
Tokens/sec129.6
Time to First Token26.24s
Time to Answer26.24s
Provider Price Ranking
Provider Price Ranking
27 providers
Cheapest: DeepInfraMost Expensive: Venice AI
ProviderInputOutput
1DeepInfraCheapest
$0
$0.00001
2Google
$0
$0.00002
3Kilo Gateway
$1
$6
4302.AI
$2
$12
5NanoGPT
$2
$12
6Abacus
$2
$12
7Perplexity Agent
$2
$12
8OpenRouter
$2
$12
9ZenMux
$2
$12
10Vivgrid
$2
$12
11FrogBot
$2
$12
12AIHubMix
$2
$12
13Requesty
$2
$12
14Vercel AI Gateway
$2
$12
15DevPass (LLM Gateway)
$2
$12
16Vertex
$2
$12
17FastRouter
$2
$12
18Auriko
$2
$12
19OrcaRouter
$2
$12
20Merge Gateway
$2
$12
21DaoXE
$2
$12
22Ofox
$2
$12
23Impossibl
$2
$12
24Eden AI
$2
$12
25Opper
$2
$12
26Tempr Gateway
$2
$12
27Venice AI
$2.5
$15
Compare pricing across different API providers for this model.