Grok 3
SpaceXAIGrokProprietary
描述
Grok 3, launched by xAI on February 17, 2025, is an advanced AI model with significantly enhanced capabilities compared to Grok 2, boasting an order of magnitude increase in performance. Trained on a vast dataset that includes legal documents among others, and utilizing a massive compute infrastructure with around 200,000 GPUs in a Memphis data center, Grok 3's training used ten times more compute than its predecessor. It features specialized models like Grok 3 Reasoning and Grok 3 Mini Reasoning for complex problem-solving, and it excels in benchmarks like AIME for mathematics and GPQA for PhD-level science.
发布日期
2025-02-19
参数规模
—
上下文长度
—
支持模态
image, text
能力雷达图
35
general
41
coding
57
reasoning
45
science
52
agents
80
multimodal
排行榜排名
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
84.6%自报
Code
LiveCodeBench
79.4%自报
General
MMMU
78.0%自报
Math
AIME 2025
93.3%自报
AIME 2024
93.3%自报
AA 评测指数
(Artificial Analysis)Math Index(Artificial Analysis)58.0
Intelligence Index(Artificial Analysis)18.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.7
Aime 25(MAA (Mathematical Association of America))0.6
Lcr(Artificial Analysis)0.6
Tau2(Sierra + U Toronto + Vector Institute (2025))0.5
Ifbench(Google Research (2023))0.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Aime(MAA (Mathematical Association of America))0.3
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
LLM Stats 分类评分
(LLM Stats (zeroeval))Math90
Reasoning90
General90
Multimodal80
Physics80
Healthcare80
Biology80
Chemistry80
Code80
Vision80
定价
输入价格$4 / 1M tokens
输出价格$20 / 1M tokens
混合价格(3:1)$8 / 1M tokens
速度
Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s
供应商价格排行
供应商价格排行
3 个供应商
最便宜: xAI最贵: SpaceXAI
供应商输入输出
1xAI最便宜
$0
$0.00002
2Helicone
$3
$15
3SpaceXAI主要
$4
$20
比较该模型在不同 API 供应商之间的定价。