跳转到主要内容

Grok-2

xAIGrokProprietary

描述

Grok-2 is a frontier language model with state-of-the-art reasoning capabilities, featuring advanced abilities in chat, coding, and reasoning. It demonstrates superior performance in visual math reasoning, document-based question answering, and excels across various academic benchmarks including reasoning, reading comprehension, math, and science.

发布日期
2024-08-13
参数规模
上下文长度
支持模态
image, text

能力雷达图

80
general
90
coding
80
reasoning
51
science估算
83
agents
90
multimodal

缺少专门科学评测时,Science 由 LLM Stats 科学得分或推理能力估算。

排行榜排名

暂无排名数据

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)56.0%自报

Code

HumanEvalOpenAI (2021)88.4%自报

Finance

MMLU87.5%自报
MMLU-Pro75.5%自报

General

MMMU66.1%自报

Image To Text

DocVQADocVQA (2020)93.6%自报

Math

MATH76.1%自报
MathVista69.0%自报

AA 评测指数

(Artificial Analysis)

暂无 AA 评测数据

LLM Stats 分类评分

(LLM Stats (zeroeval))
Image To Text
90
Code
90
Legal
80
Math
80
Multimodal
80
Language
80
Finance
80
General
80
Healthcare
80
Vision
80
Reasoning
70
Physics
60
Biology
60
Chemistry
60

定价

暂无定价数据

速度

暂无速度数据

供应商价格排行

暂无提供商数据

外部链接