Skip to main content

Gemini 3.1 Pro Preview

GoogleGeminiProprietary

Description

Gemini 3.1 Pro is the latest model in the Gemini 3 series. It excels at complex tasks requiring broad world knowledge and advanced reasoning across modalities. Gemini 3.1 Pro uses dynamic thinking by default to reason through prompts, and features a 1 million-token input context window with 64k output tokens.

Release Date
2026-02-19
Parameters
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

48
general
67
coding
94
reasoning
72
science
80
agents
80
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability70
46.0
LS
Code Ranking34
88.0
AA
General Ranking26
84.0
AA
Math Reasoning13
91.0
LB
Reasoning18
84.0
LB
Science6
94.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench99.3%SR
BrowseCompOpenAI (2025)85.9%SR
MCP Atlas69.2%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)68.5%SR
SWE-Bench ProPrinceton NLP (2024)54.2%SR
Finance Agent v243.0%
FrontierSWE40.0%
APEX-Agents33.5%SR
DeepSWE 1.112.0%
Legal Agent Benchmark0.0%

Biology

GPQANYU + Cohere + Anthropic (2023)94.3%SR
SciCode59.0%SR

Code

LiveCodeBench Pro2887.00 / 3000SR
SWE-Bench Verified80.6%SR

General

MMMLU92.6%SR
MMMU-Pro80.5%SR
LiveBench79.9%
MRCR v2 (8-needle)26.3%SR

Math

Humanity's Last Exam51.4%SR

Reasoning

ARC-AGI v277.1%SR

AA Evaluation Indices

(Artificial Analysis)
Coding Index(Artificial Analysis)
68.8
Intelligence Index(Artificial Analysis)
47.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
1.0
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Lcr(Artificial Analysis)
0.8
Ifbench(Google Research (2023))
0.8
Terminalbench V2 1
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.6
Terminalbench Hard(Stanford × Laude Institute (2026))
0.5
Hle(Center for AI Safety + Scale AI (2025))
0.5
Tau Banking
0.2

LLM Stats Category Scores

(LLM Stats (zeroeval))
Code
100
Reasoning
100
General
100
Search
90
Language
90
Multimodal
80
Physics
80
Spatial Reasoning
80
Frontend Development
80
Biology
80
Chemistry
80
Tool Calling
80
Math
70
Vision
70
Agents
50
Finance
40
Long Context
30
Healthcare
20
Legal
0

Pricing

Input Price$2 / 1M tokens
Output Price$12 / 1M tokens
Blended Price (3:1)$4.5 / 1M tokens
Cache Read Price$0.2 / 1M tokens

Speed

Tokens/sec127.4
Time to First Token17.99s
Time to Answer17.99s

Provider Price Ranking

Provider Price Ranking

22 providers

Cheapest: GoogleMost Expensive: OrcaRouter
ProviderInputOutput
1GoogleCheapest
$0
$0.00002
2NanoGPT
$2
$12
3Abacus
$2
$12
4Perplexity Agent
$2
$12
5OpenRouter
$2
$12
6ZenMux
$2
$12
7Vivgrid
$2
$12
8Kilo Gateway
$2
$12
9GitHub Copilot
$2
$12
10FrogBot
$2
$12
11AIHubMix
$2
$12
12Vercel AI Gateway
$2
$12
13LLM Gateway
$2
$12
14Vertex
$2
$12
15FastRouter
$2
$12
16Auriko
$2
$12
17Merge Gateway
$2
$12
18DaoXE
$2
$12
19Ofox
$2
$12
20Impossibl
$2
$12
21Venice AI
$2.5
$15
22OrcaRouter
$4
$18

Compare pricing across different API providers for this model.

External Sources