跳转到主要内容

Gemini 3.1 Pro Preview

GoogleGeminiProprietary

描述

Gemini 3.1 Pro is the latest model in the Gemini 3 series. It excels at complex tasks requiring broad world knowledge and advanced reasoning across modalities. Gemini 3.1 Pro uses dynamic thinking by default to reason through prompts, and features a 1 million-token input context window with 64k output tokens.

发布日期
2026-02-19
参数规模
上下文长度
1.0M
支持模态
audio, image, pdf, text, video

能力雷达图

48
general
67
coding
94
reasoning
72
science
80
agents
80
multimodal

排行榜排名

领域#排名分数来源
智能体能力模型榜70
46.0
LS
代码能力榜34
88.0
AA
通用能力榜26
84.0
AA
数学推理13
91.0
LB
推理能力18
84.0
LB
科学能力6
94.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench99.3%自报
BrowseCompOpenAI (2025)85.9%自报
MCP Atlas69.2%自报
Terminal-Bench 2.0Stanford × Laude Institute (2026)68.5%自报
SWE-Bench ProPrinceton NLP (2024)54.2%自报
Finance Agent v243.0%
FrontierSWE40.0%
APEX-Agents33.5%自报
DeepSWE 1.112.0%
Legal Agent Benchmark0.0%

Biology

GPQANYU + Cohere + Anthropic (2023)94.3%自报
SciCode59.0%自报

Code

LiveCodeBench Pro2887.00 / 3000自报
SWE-Bench Verified80.6%自报

General

MMMLU92.6%自报
MMMU-Pro80.5%自报
LiveBench79.9%
MRCR v2 (8-needle)26.3%自报

Math

Humanity's Last Exam51.4%自报

Reasoning

ARC-AGI v277.1%自报

AA 评测指数

(Artificial Analysis)
Coding Index(Artificial Analysis)
68.8
Intelligence Index(Artificial Analysis)
47.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
1.0
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Lcr(Artificial Analysis)
0.8
Ifbench(Google Research (2023))
0.8
Terminalbench V2 1
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.6
Terminalbench Hard(Stanford × Laude Institute (2026))
0.5
Hle(Center for AI Safety + Scale AI (2025))
0.5
Tau Banking
0.2

LLM Stats 分类评分

(LLM Stats (zeroeval))
Code
100
Reasoning
100
General
100
Search
90
Language
90
Multimodal
80
Physics
80
Spatial Reasoning
80
Frontend Development
80
Biology
80
Chemistry
80
Tool Calling
80
Math
70
Vision
70
Agents
50
Finance
40
Long Context
30
Healthcare
20
Legal
0

定价

输入价格$2 / 1M tokens
输出价格$12 / 1M tokens
混合价格(3:1)$4.5 / 1M tokens
缓存读取价格$0.2 / 1M tokens

速度

Tokens/秒127.4
首Token延迟17.99s
首回答延迟17.99s

供应商价格排行

供应商价格排行

22 个供应商

最便宜: Google最贵: OrcaRouter
供应商输入输出
1Google最便宜
$0
$0.00002
2NanoGPT
$2
$12
3Abacus
$2
$12
4Perplexity Agent
$2
$12
5OpenRouter
$2
$12
6ZenMux
$2
$12
7Vivgrid
$2
$12
8Kilo Gateway
$2
$12
9GitHub Copilot
$2
$12
10FrogBot
$2
$12
11AIHubMix
$2
$12
12Vercel AI Gateway
$2
$12
13LLM Gateway
$2
$12
14Vertex
$2
$12
15FastRouter
$2
$12
16Auriko
$2
$12
17Merge Gateway
$2
$12
18DaoXE
$2
$12
19Ofox
$2
$12
20Impossibl
$2
$12
21Venice AI
$2.5
$15
22OrcaRouter
$4
$18

比较该模型在不同 API 供应商之间的定价。

外部链接