Grok-2
xAIGrokProprietary
Description
Grok-2 is a frontier language model with state-of-the-art reasoning capabilities, featuring advanced abilities in chat, coding, and reasoning. It demonstrates superior performance in visual math reasoning, document-based question answering, and excels across various academic benchmarks including reasoning, reading comprehension, math, and science.
Release Date
2024-08-13
Parameters
—
Context Length
—
Modalities
image, text
Capability Radar
80
general
90
coding
80
reasoning
51
scienceest.
83
agents
90
multimodal
Science is estimated from LLM Stats science scores or reasoning when no dedicated science benchmarks are available.
Rankings
No ranking data available
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
56.0%SR
Code
HumanEvalOpenAI (2021)
88.4%SR
Finance
MMLU
87.5%SR
MMLU-Pro
75.5%SR
General
MMMU
66.1%SR
Image To Text
DocVQADocVQA (2020)
93.6%SR
Math
MATH
76.1%SR
MathVista
69.0%SR
AA Evaluation Indices
(Artificial Analysis)No AA evaluation data available
LLM Stats Category Scores
(LLM Stats (zeroeval))Image To Text90
Code90
Legal80
Math80
Multimodal80
Language80
Finance80
General80
Healthcare80
Vision80
Reasoning70
Physics60
Biology60
Chemistry60
Pricing
No pricing data available
Speed
No speed data available
Provider Price Ranking
No provider data available