Skip to main content

Gemini 2.5 Pro

GoogleGeminiProprietary

Description

A highly capable AI model from Google, designed for the agentic era. Gemini 2.5 Pro performs well on common benchmarks with enhanced reasoning, multimodal capabilities (text, image, video, audio input), and a 1M token context window.

Release Date
2025-06-05
Parameters
—
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

38
general
51
coding
89
reasoning
59
science
78
agents
80
multimodal

Rankings

Domain#RankScoreSource
Audio66
41.0
AA
Code Ranking256
57.0
AA
General Ranking235
50.0
AA
Multimodal Ranking31
60.0
LS
Science180
62.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Factuality

SimpleQA54.0%SR

General

Global-MMLU-Lite89.2%SR
Aider-Polyglot82.2%SR
Aider-Polyglot Edit72.7%SR

Long Context

MRCR93.0%SR
MRCR 1M (pointwise)82.9%SR
MRCR v2 (8-needle)16.4%SR

Math

AIME 202492.0%SR
AIME 202588.0%SR

Multimodal

Video-MME84.8%SR
VideoMMMU83.6%SR
MMMU82.0%SR
Vibe-Eval67.2%SR

Reasoning

FACTS Grounding87.8%SR
GPQANYU + Cohere + Anthropic (2023)86.4%SR
LiveCodeBench v575.6%SR
LiveCodeBench69.0%SR
SWE-Bench Verified67.2%SR
Humanity's Last Exam21.6%SR
ARC-AGI v24.9%

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
96.7
Aime(MAA (Mathematical Association of America))
88.7
Math Index(Artificial Analysis)
87.7
Aime 25(MAA (Mathematical Association of America))
87.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
86.2
Gpqa(NYU + Cohere + Anthropic (2023))
84.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))
80.1
Lcr(Artificial Analysis)
69.0
Tau2(Sierra + U Toronto + Vector Institute (2025))
54.1
Ifbench(Google Research (2023))
48.7
Scicode(UIUC + Argonne National Lab (2024))
46.3
Coding Index(Artificial Analysis)
33.3
Terminalbench V2 1
28.5
Terminalbench Hard(Stanford × Laude Institute (2026))
26.5
Hle(Center for AI Safety + Scale AI (2025))
22.5
Intelligence Index(Artificial Analysis)
16.1
Tau Banking
9.7
Terminalbench V4 0
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Language
90
Long Context
90
Multimodal
80
Physics
80
Healthcare
80
Biology
80
Chemistry
80
Reasoning
70
General
70
Code
70
Math
60
Frontend Development
60
Factuality
50
Vision
50
Spatial Reasoning
0

Pricing

Input Price$1.25 / 1M tokens
Output Price$10 / 1M tokens
Blended Price (3:1)$3.438 / 1M tokens
Cache Read Price$0.125 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

27 providers

Cheapest: GoogleMost Expensive: Cortecs
ProviderInputOutput
1GoogleCheapest
$0
$0.00001
2DeepInfra
$0
$0.00001
3Poe
$0.87
$7
4Jiekou.AI
$1.125
$9
5302.AI
$1.25
$10
6NanoGPT
$1.25
$10
7Abacus
$1.25
$10
8Perplexity Agent
$1.25
$10
9OpenRouter
$1.25
$10
10ZenMux
$1.25
$10
11Kilo Gateway
$1.25
$10
12SAP AI Core
$1.25
$10
13Helicone
$1.25
$10
14FrogBot
$1.25
$10
15AIHubMix
$1.25
$10
16Vercel AI Gateway
$1.25
$10
17DevPass (LLM Gateway)
$1.25
$10
18Vertex
$1.25
$10
19FastRouter
$1.25
$10
20Auriko
$1.25
$10
21NEAR AI Cloud
$1.25
$10
22OrcaRouter
$1.25
$10
23Merge Gateway
$1.25
$10
24Ofox
$1.25
$10
25Modelis
$1.25
$10
26Impossibl
$1.25
$10
27Cortecs
$1.495
$9.964

Compare pricing across different API providers for this model.

External Sources