Skip to main content

Hermes 4 - Llama-3.1 405B (Reasoning)

Nous ResearchLlama
Release Date
2025-08-27
Parameters
—
Context Length
131K
Modalities
text

Capability Radar

31
general
69
coding
70
reasoning
54
science
70
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking370
39.0
AA
General Ranking439
32.0
AA
Science316
45.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
82.9
Gpqa(NYU + Cohere + Anthropic (2023))
72.7
Math Index(Artificial Analysis)
69.7
Aime 25(MAA (Mathematical Association of America))
69.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
68.6
Ifbench(Google Research (2023))
32.7
Lcr(Artificial Analysis)
22.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
22.2
Terminalbench Hard(Stanford × Laude Institute (2026))
11.4
Hle(Center for AI Safety + Scale AI (2025))
10.9
Intelligence Index(Artificial Analysis)
7.5

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input Price$1 / 1M tokens
Output Price$3 / 1M tokens
Blended Price (3:1)$1.5 / 1M tokens

Speed

Tokens/sec43.5
Time to First Token0.74s
Time to Answer46.68s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1Nous ResearchPRIMARY
$1
$3

Compare pricing across different API providers for this model.

External Sources