Gemini 1.5 Pro (Sep '24)
GoogleGeminiProprietary
描述
Gemini 1.5 Pro is a mid-size multimodal model optimized for a wide range of reasoning tasks. It can process large amounts of data at once, including 2 hours of video, 19 hours of audio, codebases with 60,000 lines of code, or 2,000 pages of text.
發布日期
2024-09-24
參數規模
—
上下文長度
—
支援模態
image, text
能力雷達圖
28
general
27
coding
50
reasoning
42
science
43
agents
80
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))General
MMLU
85.9%自報
Language
FLEURS
93.3%自報
MMLU-Pro
75.8%自報
WMT23
75.1%自報
Long Context
MRCR
82.6%自報
Math
GSM8k
90.8%自報
MGSM
87.5%自報
MATH
86.5%自報
MathVista
68.1%自報
FunctionalMATH
64.6%自報
HiddenMath
52.0%自報
AMC_2022_23
46.4%自報
Multimodal
Video-MME
78.6%自報
MMMU
65.9%自報
Vibe-Eval
53.9%自報
Physics
PhysicsFinals
63.9%自報
Reasoning
HellaSwagAI2 (2019)
93.3%自報
BIG-Bench Hard
89.2%自報
Natural2Code
85.4%自報
HumanEvalOpenAI (2021)
84.1%自報
DROP
74.9%自報
GPQANYU + Cohere + Anthropic (2023)
59.1%自報
Safety
XSTest
98.8%自報
AA 評測指數
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))87.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))75.0
Gpqa(NYU + Cohere + Anthropic (2023))58.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))31.6
Coding Index(Artificial Analysis)23.6
Aime(MAA (Mathematical Association of America))23.0
Intelligence Index(Artificial Analysis)7.9
Hle(Center for AI Safety + Scale AI (2025))4.6
LLM Stats 分類評分
(LLM Stats (zeroeval))Safety100
Speech To Text90
Language80
Legal80
Long Context80
Math80
Reasoning80
Finance80
General80
Healthcare80
Code80
Multimodal70
Vision70
Physics60
Biology60
Chemistry60
定價
輸入價格免費
輸出價格免費
混合價格(3:1)免費
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
暫無提供商資料