Grok-2
xAIGrokProprietary
描述
Grok-2 is a frontier language model with state-of-the-art reasoning capabilities, featuring advanced abilities in chat, coding, and reasoning. It demonstrates superior performance in visual math reasoning, document-based question answering, and excels across various academic benchmarks including reasoning, reading comprehension, math, and science.
發布日期
2024-08-13
參數規模
—
上下文長度
—
支援模態
image, text
能力雷達圖
80
general
90
coding
80
reasoning
51
science估算
83
agents
90
multimodal
缺少專門科學評測時,Science 由 LLM Stats 科學得分或推理能力估算。
排行榜排名
| 領域 | #排名 | 分數 | 來源 |
|---|---|---|---|
| 多模態榜 | 97 | 47.0 | LS |
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))General
MMLU
87.5%自報
Language
MMLU-Pro
75.5%自報
Math
MATH
76.1%自報
MathVista
69.0%自報
Multimodal
MMMU
66.1%自報
Reasoning
HumanEvalOpenAI (2021)
88.4%自報
GPQANYU + Cohere + Anthropic (2023)
56.0%自報
Vision
DocVQADocVQA (2020)
93.6%自報
AA 評測指數
(Artificial Analysis)暫無 AA 評測資料
LLM Stats 分類評分
(LLM Stats (zeroeval))Image To Text90
Code90
Language80
Legal80
Math80
Multimodal80
Finance80
General80
Healthcare80
Vision80
Reasoning70
Physics60
Biology60
Chemistry60
定價
暫無定價資料
速度
暫無速度資料
供應商價格排行
暫無提供商資料