跳轉到主要內容

Gemini 2.0 Flash (Feb '25)

GoogleGeminiProprietary

描述

Next-generation model featuring superior speed, native tool use, multimodal generation, and a 1M token context window. Supports audio, images, video, and text input with capabilities for structured outputs, function calling, code execution, search, and multimodal operations.

發布日期
2025-02-05
參數規模
上下文長度
支援模態
image, text

能力雷達圖

31
general
33
coding
39
reasoning
41
science
37
agents
80
multimodal

排行榜排名

領域#排名分數來源
程式碼能力榜373
27.0
AA
通用能力榜341
38.0
AA
科學能力339
41.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Audio

CoVoST239.2%自報

Biology

GPQANYU + Cohere + Anthropic (2023)74.2%自報

Code

LiveCodeBench35.1%自報

Factuality

FACTS Grounding83.6%自報

Finance

MMLU-Pro76.4%自報

General

Natural2Code92.9%自報
MMMU75.4%自報
MRCR69.2%自報
Vibe-Eval56.3%自報

Long Context

EgoSchema71.5%自報

Math

MATH89.7%自報
AIME 202473.3%自報
HiddenMath63.0%自報

Reasoning

Bird-SQL (dev)56.9%自報

AA 評測指數

(Artificial Analysis)
Math Index(Artificial Analysis)
21.7
Intelligence Index(Artificial Analysis)
12.2
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.6
Ifbench(Google Research (2023))
0.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.3
Scicode(UIUC + Argonne National Lab (2024))
0.3
Aime(MAA (Mathematical Association of America))
0.3
Lcr(Artificial Analysis)
0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.3
Aime 25(MAA (Mathematical Association of America))
0.2
Hle(Center for AI Safety + Scale AI (2025))
0.0
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Stats 分類評分

(LLM Stats (zeroeval))
Legal
80
Math
80
Factuality
80
Finance
80
Grounding
80
Long Context
70
Reasoning
70
General
70
Healthcare
70
Vision
70
Multimodal
60
Physics
60
Language
60
Biology
60
Chemistry
60
Speech To Text
40
Audio
40
Code
40

定價

輸入價格免費
輸出價格免費
混合價格(3:1)免費

速度

Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s

供應商價格排行

暫無提供商資料

外部連結