Skip to main content

Hermes 4 - Llama-3.1 70B (Reasoning)

Nous ResearchLlama
Release Date
2025-08-27
Parameters
Context Length
131K
Modalities
text

Capability Radar

31
general
58
coding
69
reasoning
45
science
66
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking356
29.0
AA
General Ranking376
34.0
AA
Science273
47.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Math Index(Artificial Analysis)
68.7
Intelligence Index(Artificial Analysis)
9.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.7
Aime 25(MAA (Mathematical Association of America))
0.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.3
Ifbench(Google Research (2023))
0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.2
Hle(Center for AI Safety + Scale AI (2025))
0.1
Lcr(Artificial Analysis)
0.1
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input Price$0.13 / 1M tokens
Output Price$0.4 / 1M tokens
Blended Price (3:1)$0.198 / 1M tokens

Speed

Tokens/sec88.8
Time to First Token0.58s
Time to Answer23.09s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1Nous ResearchPRIMARY
$0.13
$0.4

Compare pricing across different API providers for this model.

External Sources