Grok 4.5 (high)
SpaceXAIGrokProprietary
描述
Grok 4.5 is SpaceXAI's frontier model for coding, agentic tasks, and knowledge work. Trained alongside Cursor across tens of thousands of NVIDIA GB300 GPUs, it uses large-scale reinforcement learning focused on multi-step software engineering and other technical work. The model supports text and image inputs, reasoning, function calling, structured outputs, and a 500K-token context window, and is served at 80 tokens per second.
發布日期
2026-07-08
參數規模
—
上下文長度
500K
支援模態
image, pdf, text
能力雷達圖
53
general
70
coding
93
reasoning
69
science
60
agents
80
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Agents
AutomationBench-AA
51.4%
Tau3 Banking
33.0%
Code
DeepSWE 1.0
62.0%自報
DeepSWE 1.1
54.0%
DeepSWE
53.0%自報
SWE-Marathon
29.0%自報
General
Artificial Analysis
54.0%
Knowledge
AA-Omniscience Index
126.00 / 200
OmniScience (non-hallucination rate)
46.0%
Reasoning
GPQANYU + Cohere + Anthropic (2023)
93.0%
Terminal-Bench 2.1
83.3%自報
SWE-Bench ProPrinceton NLP (2024)
64.7%自報
FrontierCode 1.1
42.4%
Science
OmniScience
52.0%
AA 評測指數
(Artificial Analysis)Coding Index(Artificial Analysis)72.4
Intelligence Index(Artificial Analysis)55.8
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Terminalbench V2 10.8
Lcr(Artificial Analysis)0.7
Scicode(UIUC + Argonne National Lab (2024))0.5
Hle(Center for AI Safety + Scale AI (2025))0.4
Tau Banking0.4
LLM Stats 分類評分
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Reasoning60
General60
Tool Calling60
Knowledge50
Science50
Agents50
Code50
定價
輸入價格$2 / 1M tokens
輸出價格$6 / 1M tokens
混合價格(3:1)$3 / 1M tokens
快取讀取價格$0.3 / 1M tokens
速度
Tokens/秒55.3
首Token延遲9.67s
首回答延遲9.67s
供應商價格排行
供應商價格排行
18 個供應商
最便宜: xAI最貴: Venice AI
供應商輸入輸出
1xAI最便宜
$0
$0.00001
2Requesty
$1.8
$5.4
3SpaceXAI主要
$2
$6
4NanoGPT
$2
$6
5Abacus
$2
$6
6OpenRouter
$2
$6
7OpenCode Go
$2
$6
8ZenMux
$2
$6
9Kilo Gateway
$2
$6
10GitHub Copilot
$2
$6
11OpenCode Zen
$2
$6
12AIHubMix
$2
$6
13DevPass (LLM Gateway)
$2
$6
14CrossModel
$2
$6
15Pioneer
$2
$6
16DaoXE
$2
$6
17Ofox
$2
$6
18Venice AI
$2.27
$6.8
比較該模型在不同 API 供應商之間的定價。