跳轉到主要內容

Hermes 4 - Llama-3.1 70B (Reasoning)

Nous ResearchLlama
發布日期
2025-08-27
參數規模
—
上下文長度
131K
支援模態
text

能力雷達圖

30
general
65
coding
69
reasoning
51
science
68
agents
0
multimodal

排行榜排名

領域#排名分數來源
程式碼能力榜435
30.0
AA
通用能力榜439
32.0
AA
科學能力371
41.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

暫無基準測試資料

AA 評測指數

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
81.1
Gpqa(NYU + Cohere + Anthropic (2023))
69.9
Math Index(Artificial Analysis)
68.7
Aime 25(MAA (Mathematical Association of America))
68.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
65.3
Ifbench(Google Research (2023))
31.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
22.5
Lcr(Artificial Analysis)
9.7
Hle(Center for AI Safety + Scale AI (2025))
8.8
Intelligence Index(Artificial Analysis)
7.9
Terminalbench Hard(Stanford × Laude Institute (2026))
4.5

LLM Stats 分類評分

(LLM Stats (zeroeval))

暫無分類評分資料

定價

輸入價格免費
輸出價格免費
混合價格(3:1)免費

速度

Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s

供應商價格排行

暫無提供商資料

外部連結