跳转到主要内容

Gemini 3.7 Flash (high)

GoogleGeminiProprietary

描述

Google's most intelligent Flash workhorse yet for coding and agents. Ships ~3 weeks after Gemini 3.6 Flash with gains on debugging/issue resolution, first-pass code accuracy, FrontierCode 1.1 Main (43.6% vs 3.6's 34.4%), and DeepSWE v1.1 (65.3% vs 49.0%). WebDev Arena Elo 1588 vs 1538; GDP.pdf 34.0% vs 22.0%; AutomationBench 30.4% vs 17.0%. Text/image/video/audio/PDF in, text out; 1M context / 65,536 max output. Powers Gemini Spark for Pro/Ultra. Intro pricing is half of original 3.6 Flash ($0.75/$3.75 per 1M input/output tokens) through 2026-12-31; list price $1.50/$7.50 applies starting 2027-01-01.

发布日期
2026-08-13
参数规模
上下文长度
1.0M
支持模态
audio, image, pdf, text, video

能力雷达图

54
general
73
coding
95
reasoning
72
science
50
agents
80
multimodal

排行榜排名

领域#排名分数来源
智能体能力模型榜32
54.0
LS
代码能力榜3
97.0
AA
通用能力榜13
89.0
AA
多模态榜17
62.0
LS
科学能力7
93.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Agents

GDPval-AA1525.00 / 3000自报
Terminal-Bench 2.185.8%自报
DeepSWE 1.165.3%自报
OSWorld 2.047.9%自报
FrontierCode 1.143.6%自报
AutomationBench30.4%自报
Agents' Last Exam26.3%自报
Terminal-Bench 3.014.9%自报

Biology

BioMysteryBench43.5%自报

Code

WebDev Arena1588.00 / 2000自报

General

MRCR v2 (8-needle)97.0%自报
Artificial Analysis56.0%自报
GDP.pdf34.0%自报

Long Context

LVBench85.4%自报

Multimodal

CharXiv-R88.7%自报

AA 评测指数

(Artificial Analysis)
Coding Index(Artificial Analysis)
76.1
Intelligence Index(Artificial Analysis)
56.0
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Terminalbench V2 1
0.9
Lcr(Artificial Analysis)
0.8
Scicode(UIUC + Argonne National Lab (2024))
0.6
Hle(Center for AI Safety + Scale AI (2025))
0.5
Tau Banking
0.3

LLM Stats 分类评分

(LLM Stats (zeroeval))
Legal
100
Finance
100
Agents
100
Reasoning
100
General
100
Long Context
90
Multimodal
60
Code
60
Vision
60
Tool Calling
50
Science
40
Biology
40

定价

输入价格$0.75 / 1M tokens
输出价格$3.75 / 1M tokens
混合价格(3:1)$1.5 / 1M tokens
缓存读取价格$0.075 / 1M tokens

速度

Tokens/秒515.0
首Token延迟6.35s
首回答延迟6.35s

供应商价格排行

供应商价格排行

1 个供应商

供应商输入输出
1Google
$0
$0

比较该模型在不同 API 供应商之间的定价。

外部链接