メインコンテンツへスキップ

Gemini 2.5 Flash Preview (Reasoning)

GoogleGeminiProprietary

説明

A thinking model designed for a balance between price and performance. It builds upon Gemini 2.0 Flash with upgraded reasoning, hybrid thinking control, multimodal capabilities (text, image, video, audio input), and a 1M token input context window.

リリース日
2025-04-17
パラメータ
—
コンテキスト長
1.0M
モダリティ
audio, image, pdf, text, video

能力レーダー

32
general
50
coding
86
reasoning
52
science
75
agents
80
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング271
55.0
AA
総合ランキング344
40.0
AA
マルチモーダルランキング45
58.0
LS
科学326
44.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Factuality

SimpleQA26.9%自己申告

General

Global-MMLU-Lite88.4%自己申告
Aider-Polyglot61.9%自己申告
Aider-Polyglot Edit56.7%自己申告

Long Context

MRCR32.0%自己申告

Math

AIME 202488.0%自己申告
AIME 202572.0%自己申告

Multimodal

MMMU79.7%自己申告
Vibe-Eval65.4%自己申告

Reasoning

FACTS Grounding85.3%自己申告
GPQANYU + Cohere + Anthropic (2023)82.8%自己申告
LiveCodeBench v563.9%自己申告
SWE-Bench Verified60.4%自己申告
Humanity's Last Exam11.0%自己申告

AA評価指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
98.1
Aime(MAA (Mathematical Association of America))
84.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
80.0
Gpqa(NYU + Cohere + Anthropic (2023))
69.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))
50.5
Hle(Center for AI Safety + Scale AI (2025))
12.1
Intelligence Index(Artificial Analysis)
11.7

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Language
90
Grounding
90
Physics
80
Healthcare
80
Biology
80
Chemistry
80
Multimodal
70
Math
60
Reasoning
60
Factuality
60
Frontend Development
60
General
60
Code
60
Vision
50
Long Context
20

価格設定

入力価格無料
出力価格無料
混合価格(3:1)無料
キャッシュ読み取り価格$0.03 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

1 プロバイダー

プロバイダー入力出力
1DeepInfra
$0
$0

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク