Gemini 3 Pro Preview (low)
GoogleGeminiProprietary
Description
Gemini 3 Pro is the first model in the new Gemini 3 series. It is best for complex tasks that require broad world knowledge and advanced reasoning across modalities. Gemini 3 Pro uses dynamic thinking by default to reason through prompts, and features a 1 million-token input context window with 64k output tokens.
Release Date
2025-11-18
Parameters
—
Context Length
—
Modalities
audio, image, text, video
Capability Radar
44
general
86
coding
87
reasoning
70
science
70
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 26 | 57.0 | LS |
| Code Ranking | 137 | 76.0 | AA |
| General Ranking | 169 | 58.0 | AA |
| Multimodal Ranking | 36 | 60.0 | LS |
| Science | 129 | 69.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
85.4%SR
Factuality
SimpleQA
72.1%SR
Language
MMMLU
91.8%SR
Long Context
MRCR v2 (8-needle)
26.3%SR
Math
AIME 2025
100.0%SR
LiveBench
73.4%
MathArena Apex
23.4%SR
Multimodal
VideoMMMU
87.6%SR
Reasoning
Vending-Bench 2
547816.0%SR
LiveCodeBench Pro
2439.00 / 3000SR
Global PIQA
93.4%SR
GPQANYU + Cohere + Anthropic (2023)
91.9%SR
CharXiv-R
81.4%SR
SWE-Bench Verified
76.2%SR
FACTS Grounding
70.5%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)
54.2%SR
Humanity's Last Exam
45.8%SR
ARC-AGI v2
31.1%SR
Vision
MMMU-Pro
81.0%SR
ScreenSpot Pro
72.7%SR
OmniDocBench 1.5
11.5%SR
AA Evaluation Indices
(Artificial Analysis)Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))89.5
Gpqa(NYU + Cohere + Anthropic (2023))88.7
Math Index(Artificial Analysis)86.7
Aime 25(MAA (Mathematical Association of America))86.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))85.7
Lcr(Artificial Analysis)74.0
Tau2(Sierra + U Toronto + Vector Institute (2025))68.1
Ifbench(Google Research (2023))49.7
Terminalbench Hard(Stanford × Laude Institute (2026))34.1
Hle(Center for AI Safety + Scale AI (2025))29.5
Intelligence Index(Artificial Analysis)22.3
LLM Stats Category Scores
(LLM Stats (zeroeval))Agents100
Code100
Reasoning100
General100
Language90
Physics90
Healthcare90
Biology90
Chemistry90
Frontend Development80
Math70
Multimodal70
Factuality70
Grounding70
Tool Calling70
Vision60
Spatial Reasoning50
Long Context30
Structured Output10
Pricing
Input Price$2 / 1M tokens
Output Price$12 / 1M tokens
Blended Price (3:1)$4.5 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1GooglePRIMARY
$2
$12
Compare pricing across different API providers for this model.