Gemini 2.5 Flash Preview (Reasoning)
GoogleGeminiProprietary
説明
A thinking model designed for a balance between price and performance. It builds upon Gemini 2.0 Flash with upgraded reasoning, hybrid thinking control, multimodal capabilities (text, image, video, audio input), and a 1M token input context window.
リリース日
2025-04-17
パラメータ
—
コンテキスト長
1.0M
モダリティ
audio, image, pdf, text, video
能力レーダー
32
general
50
coding
86
reasoning
52
science
75
agents
80
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 271 | 55.0 | AA |
| 総合ランキング | 344 | 40.0 | AA |
| マルチモーダルランキング | 45 | 58.0 | LS |
| 科学 | 326 | 44.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Factuality
SimpleQA
26.9%自己申告
General
Global-MMLU-Lite
88.4%自己申告
Aider-Polyglot
61.9%自己申告
Aider-Polyglot Edit
56.7%自己申告
Long Context
MRCR
32.0%自己申告
Math
AIME 2024
88.0%自己申告
AIME 2025
72.0%自己申告
Multimodal
MMMU
79.7%自己申告
Vibe-Eval
65.4%自己申告
Reasoning
FACTS Grounding
85.3%自己申告
GPQANYU + Cohere + Anthropic (2023)
82.8%自己申告
LiveCodeBench v5
63.9%自己申告
SWE-Bench Verified
60.4%自己申告
Humanity's Last Exam
11.0%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))98.1
Aime(MAA (Mathematical Association of America))84.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))80.0
Gpqa(NYU + Cohere + Anthropic (2023))69.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))50.5
Hle(Center for AI Safety + Scale AI (2025))12.1
Intelligence Index(Artificial Analysis)11.7
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Language90
Grounding90
Physics80
Healthcare80
Biology80
Chemistry80
Multimodal70
Math60
Reasoning60
Factuality60
Frontend Development60
General60
Code60
Vision50
Long Context20
価格設定
入力価格無料
出力価格無料
混合価格(3:1)無料
キャッシュ読み取り価格$0.03 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
1 プロバイダー
プロバイダー入力出力
1DeepInfra
$0
$0
このモデルの異なるAPIプロバイダー間の価格を比較。