Hermes 4 - Llama-3.1 70B (Reasoning)
Nous ResearchLlama
發布日期
2025-08-27
參數規模
—
上下文長度
131K
支援模態
text
能力雷達圖
30
general
65
coding
69
reasoning
51
science
68
agents
0
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))暫無基準測試資料
AA 評測指數
(Artificial Analysis)Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))81.1
Gpqa(NYU + Cohere + Anthropic (2023))69.9
Math Index(Artificial Analysis)68.7
Aime 25(MAA (Mathematical Association of America))68.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))65.3
Ifbench(Google Research (2023))31.3
Tau2(Sierra + U Toronto + Vector Institute (2025))22.5
Lcr(Artificial Analysis)9.7
Hle(Center for AI Safety + Scale AI (2025))8.8
Intelligence Index(Artificial Analysis)7.9
Terminalbench Hard(Stanford × Laude Institute (2026))4.5
LLM Stats 分類評分
(LLM Stats (zeroeval))暫無分類評分資料
定價
輸入價格免費
輸出價格免費
混合價格(3:1)免費
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
暫無提供商資料