跳转到主要内容

Gemini 3.1 Pro Preview

GoogleGeminiProprietary

描述

Gemini 3.1 Pro is the latest model in the Gemini 3 series. It excels at complex tasks requiring broad world knowledge and advanced reasoning across modalities. Gemini 3.1 Pro uses dynamic thinking by default to reason through prompts, and features a 1 million-token input context window with 64k output tokens.

发布日期
2026-02-19
参数规模
—
上下文长度
1.0M
支持模态
audio, image, pdf, text, video

能力雷达图

33
general
67
coding
94
reasoning
72
science
80
agents
80
multimodal

排行榜排名

领域#排名分数来源
智能体能力模型榜19
60.0
LS
代码能力榜89
86.0
AA
通用能力榜68
71.0
AA
数学推理27
91.0
LB
多模态榜32
60.0
LS
推理能力34
84.0
LB
科学能力26
86.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench99.3%自报
Finance Agent v243.0%

Code

FrontierSWE40.0%
DeepSWE 1.112.0%

Language

MMMLU92.6%自报

Legal

Legal Agent Benchmark0.0%

Long Context

MRCR v2 (8-needle)26.3%自报

Math

LiveBench79.9%

Reasoning

LiveCodeBench Pro2887.00 / 3000自报
GPQANYU + Cohere + Anthropic (2023)94.3%自报
BrowseCompOpenAI (2025)85.9%自报
SWE-Bench Verified80.6%自报
ARC-AGI v277.1%自报
MCP Atlas69.2%自报
Terminal-Bench 2.0Stanford × Laude Institute (2026)68.5%自报
SciCode59.0%自报
SWE-Bench ProPrinceton NLP (2024)54.2%自报
Humanity's Last Exam51.4%自报
APEX-Agents33.5%自报

Vision

MMMU-Pro80.5%自报

AA 评测指数

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
95.6
Gpqa(NYU + Cohere + Anthropic (2023))
94.1
Lcr(Artificial Analysis)
82.0
Ifbench(Google Research (2023))
77.1
Terminalbench V2 1
73.8
Coding Index(Artificial Analysis)
68.8
Scicode(UIUC + Argonne National Lab (2024))
58.7
Terminalbench Hard(Stanford × Laude Institute (2026))
53.8
Hle(Center for AI Safety + Scale AI (2025))
47.0
Intelligence Index(Artificial Analysis)
29.7
Tau Banking
21.4
Terminalbench V4 0
4.0

LLM Stats 分类评分

(LLM Stats (zeroeval))
Code
100
Reasoning
100
General
100
Language
90
Search
90
Multimodal
80
Physics
80
Spatial Reasoning
80
Frontend Development
80
Biology
80
Chemistry
80
Tool Calling
80
Math
70
Vision
70
Agents
50
Finance
40
Long Context
30
Healthcare
20
Legal
0

定价

输入价格$2 / 1M tokens
输出价格$12 / 1M tokens
混合价格(3:1)$4.5 / 1M tokens
缓存读取价格$0.2 / 1M tokens

速度

Tokens/秒127.1
首Token延迟25.82s
首回答延迟25.82s

供应商价格排行

供应商价格排行

27 个供应商

最便宜: DeepInfra最贵: Venice AI
供应商输入输出
1DeepInfra最便宜
$0
$0.00001
2Google
$0
$0.00002
3Kilo Gateway
$1
$6
4302.AI
$2
$12
5NanoGPT
$2
$12
6Abacus
$2
$12
7Perplexity Agent
$2
$12
8OpenRouter
$2
$12
9ZenMux
$2
$12
10Vivgrid
$2
$12
11FrogBot
$2
$12
12AIHubMix
$2
$12
13Requesty
$2
$12
14Vercel AI Gateway
$2
$12
15DevPass (LLM Gateway)
$2
$12
16Vertex
$2
$12
17FastRouter
$2
$12
18Auriko
$2
$12
19OrcaRouter
$2
$12
20Merge Gateway
$2
$12
21DaoXE
$2
$12
22Ofox
$2
$12
23Impossibl
$2
$12
24Eden AI
$2
$12
25Opper
$2
$12
26Tempr Gateway
$2
$12
27Venice AI
$2.5
$15

比较该模型在不同 API 供应商之间的定价。

外部链接