K2 Think V2
MBZUAI Institute of Foundation Models
發布日期
2025-12-15
參數規模
—
上下文長度
—
支援模態
—
能力雷達圖
16
general
23
coding
71
reasoning
46
science
57
agents
0
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))暫無基準測試資料
AA 評測指數
(Artificial Analysis)Coding Index(Artificial Analysis)21.0
Intelligence Index(Artificial Analysis)17.4
Gpqa(NYU + Cohere + Anthropic (2023))0.7
Ifbench(Google Research (2023))0.6
Lcr(Artificial Analysis)0.6
Scicode(UIUC + Argonne National Lab (2024))0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))0.3
Terminalbench V2 10.1
Hle(Center for AI Safety + Scale AI (2025))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
LLM Stats 分類評分
(LLM Stats (zeroeval))暫無分類評分資料
定價
輸入價格免費
輸出價格免費
混合價格(3:1)免費
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
暫無提供商資料