Skip to main content

Grok-2

xAIGrokProprietary

Description

Grok-2 is a frontier language model with state-of-the-art reasoning capabilities, featuring advanced abilities in chat, coding, and reasoning. It demonstrates superior performance in visual math reasoning, document-based question answering, and excels across various academic benchmarks including reasoning, reading comprehension, math, and science.

Release Date
2024-08-13
Parameters
—
Context Length
—
Modalities
image, text

Capability Radar

80
general
90
coding
80
reasoning
51
scienceest.
83
agents
90
multimodal

Science is estimated from LLM Stats science scores or reasoning when no dedicated science benchmarks are available.

Rankings

Domain#RankScoreSource
Multimodal Ranking97
47.0
LS

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

General

MMLU87.5%SR

Language

MMLU-Pro75.5%SR

Math

MATH76.1%SR
MathVista69.0%SR

Multimodal

MMMU66.1%SR

Reasoning

HumanEvalOpenAI (2021)88.4%SR
GPQANYU + Cohere + Anthropic (2023)56.0%SR

Vision

DocVQADocVQA (2020)93.6%SR

AA Evaluation Indices

(Artificial Analysis)

No AA evaluation data available

LLM Stats Category Scores

(LLM Stats (zeroeval))
Image To Text
90
Code
90
Language
80
Legal
80
Math
80
Multimodal
80
Finance
80
General
80
Healthcare
80
Vision
80
Reasoning
70
Physics
60
Biology
60
Chemistry
60

Pricing

No pricing data available

Speed

No speed data available

Provider Price Ranking

No provider data available

External Sources