Grok 3
SpaceXAIGrokProprietary
描述
Grok 3, launched by xAI on February 17, 2025, is an advanced AI model with significantly enhanced capabilities compared to Grok 2, boasting an order of magnitude increase in performance. Trained on a vast dataset that includes legal documents among others, and utilizing a massive compute infrastructure with around 200,000 GPUs in a Memphis data center, Grok 3's training used ten times more compute than its predecessor. It features specialized models like Grok 3 Reasoning and Grok 3 Mini Reasoning for complex problem-solving, and it excels in benchmarks like AIME for mathematics and GPQA for PhD-level science.
發布日期
2025-02-19
參數規模
—
上下文長度
—
支援模態
image, text
能力雷達圖
35
general
41
coding
57
reasoning
45
science
52
agents
80
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
84.6%自報
Code
LiveCodeBench
79.4%自報
General
MMMU
78.0%自報
Math
AIME 2025
93.3%自報
AIME 2024
93.3%自報
AA 評測指數
(Artificial Analysis)Math Index(Artificial Analysis)58.0
Intelligence Index(Artificial Analysis)18.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.7
Aime 25(MAA (Mathematical Association of America))0.6
Lcr(Artificial Analysis)0.6
Tau2(Sierra + U Toronto + Vector Institute (2025))0.5
Ifbench(Google Research (2023))0.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Aime(MAA (Mathematical Association of America))0.3
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
LLM Stats 分類評分
(LLM Stats (zeroeval))Math90
Reasoning90
General90
Multimodal80
Physics80
Healthcare80
Biology80
Chemistry80
Code80
Vision80
定價
輸入價格$4 / 1M tokens
輸出價格$20 / 1M tokens
混合價格(3:1)$8 / 1M tokens
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
供應商價格排行
3 個供應商
最便宜: xAI最貴: SpaceXAI
供應商輸入輸出
1xAI最便宜
$0
$0.00002
2Helicone
$3
$15
3SpaceXAI主要
$4
$20
比較該模型在不同 API 供應商之間的定價。