Skip to main content

Hermes 4 - Llama-3.1 70B (Reasoning)

Nous ResearchLlama
Release Date
2025-08-27
Parameters
—
Context Length
131K
Modalities
text

Capability Radar

30
general
65
coding
69
reasoning
51
science
68
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking433
30.0
AA
General Ranking437
32.0
AA
Science369
41.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
81.1
Gpqa(NYU + Cohere + Anthropic (2023))
69.9
Math Index(Artificial Analysis)
68.7
Aime 25(MAA (Mathematical Association of America))
68.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
65.3
Ifbench(Google Research (2023))
31.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
22.5
Lcr(Artificial Analysis)
9.7
Hle(Center for AI Safety + Scale AI (2025))
8.8
Intelligence Index(Artificial Analysis)
7.9
Terminalbench Hard(Stanford × Laude Institute (2026))
4.5

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

No provider data available

External Sources