Grok 3
SpaceXAIGrokProprietary
描述
Grok 3, launched by xAI on February 17, 2025, is an advanced AI model with significantly enhanced capabilities compared to Grok 2, boasting an order of magnitude increase in performance. Trained on a vast dataset that includes legal documents among others, and utilizing a massive compute infrastructure with around 200,000 GPUs in a Memphis data center, Grok 3's training used ten times more compute than its predecessor. It features specialized models like Grok 3 Reasoning and Grok 3 Mini Reasoning for complex problem-solving, and it excels in benchmarks like AIME for mathematics and GPQA for PhD-level science.
发布日期
2025-02-19
参数规模
—
上下文长度
—
支持模态
image, text
能力雷达图
31
general
43
coding
57
reasoning
49
science
53
agents
80
multimodal
排行榜排名
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Math
AIME 2025
93.3%自报
AIME 2024
93.3%自报
Multimodal
MMMU
78.0%自报
Reasoning
GPQANYU + Cohere + Anthropic (2023)
84.6%自报
LiveCodeBench
79.4%自报
AA 评测指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))87.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))79.9
Gpqa(NYU + Cohere + Anthropic (2023))69.3
Math Index(Artificial Analysis)58.0
Aime 25(MAA (Mathematical Association of America))58.0
Lcr(Artificial Analysis)58.0
Tau2(Sierra + U Toronto + Vector Institute (2025))48.8
Ifbench(Google Research (2023))46.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))42.5
Aime(MAA (Mathematical Association of America))33.0
Intelligence Index(Artificial Analysis)12.1
Terminalbench Hard(Stanford × Laude Institute (2026))11.4
Hle(Center for AI Safety + Scale AI (2025))4.1
LLM Stats 分类评分
(LLM Stats (zeroeval))Math90
Reasoning90
General90
Multimodal80
Physics80
Healthcare80
Biology80
Chemistry80
Code80
Vision80
定价
输入价格$4 / 1M tokens
输出价格$20 / 1M tokens
混合价格(3:1)$8 / 1M tokens
速度
Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s
供应商价格排行
供应商价格排行
3 个供应商
最便宜: xAI最贵: SpaceXAI
供应商输入输出
1xAI最便宜
$0
$0.00002
2Helicone
$3
$15
3SpaceXAI主要
$4
$20
比较该模型在不同 API 供应商之间的定价。