跳转到主要内容

Hermes 4 - Llama-3.1 70B (Reasoning)

Nous ResearchLlama
发布日期
2025-08-27
参数规模
—
上下文长度
131K
支持模态
text

能力雷达图

30
general
65
coding
69
reasoning
51
science
68
agents
0
multimodal

排行榜排名

领域#排名分数来源
代码能力榜426
30.0
AA
通用能力榜432
32.0
AA
科学能力363
41.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

暂无基准测试数据

AA 评测指数

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
81.1
Gpqa(NYU + Cohere + Anthropic (2023))
69.9
Math Index(Artificial Analysis)
68.7
Aime 25(MAA (Mathematical Association of America))
68.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
65.3
Ifbench(Google Research (2023))
31.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
22.5
Lcr(Artificial Analysis)
9.7
Hle(Center for AI Safety + Scale AI (2025))
8.8
Intelligence Index(Artificial Analysis)
7.9
Terminalbench Hard(Stanford × Laude Institute (2026))
4.5

LLM Stats 分类评分

(LLM Stats (zeroeval))

暂无分类评分数据

定价

输入价格免费
输出价格免费
混合价格(3:1)免费

速度

Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s

供应商价格排行

暂无提供商数据

外部链接