Grok-2
xAIGrokProprietary
描述
Grok-2 is a frontier language model with state-of-the-art reasoning capabilities, featuring advanced abilities in chat, coding, and reasoning. It demonstrates superior performance in visual math reasoning, document-based question answering, and excels across various academic benchmarks including reasoning, reading comprehension, math, and science.
發布日期
2024-08-13
參數規模
—
上下文長度
—
支援模態
image, text
能力雷達圖
80
general
90
coding
80
reasoning
51
science估算
83
agents
90
multimodal
缺少專門科學評測時,Science 由 LLM Stats 科學得分或推理能力估算。
排行榜排名
暫無排名資料
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
56.0%自報
Code
HumanEvalOpenAI (2021)
88.4%自報
Finance
MMLU
87.5%自報
MMLU-Pro
75.5%自報
General
MMMU
66.1%自報
Image To Text
DocVQADocVQA (2020)
93.6%自報
Math
MATH
76.1%自報
MathVista
69.0%自報
AA 評測指數
(Artificial Analysis)暫無 AA 評測資料
LLM Stats 分類評分
(LLM Stats (zeroeval))Image To Text90
Code90
Legal80
Math80
Multimodal80
Language80
Finance80
General80
Healthcare80
Vision80
Reasoning70
Physics60
Biology60
Chemistry60
定價
暫無定價資料
速度
暫無速度資料
供應商價格排行
暫無提供商資料