Skip to main content

Hermes 4 - Llama-3.1 405B (Reasoning)

Nous ResearchLlama
Release Date
2025-08-27
Parameters
Context Length
131K
Modalities
text

Capability Radar

31
general
59
coding
70
reasoning
44
science
67
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking289
40.0
AA
General Ranking377
34.0
AA
Science299
45.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Math Index(Artificial Analysis)
69.7
Intelligence Index(Artificial Analysis)
8.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.7
Aime 25(MAA (Mathematical Association of America))
0.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.7
Ifbench(Google Research (2023))
0.3
Scicode(UIUC + Argonne National Lab (2024))
0.3
Lcr(Artificial Analysis)
0.2
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.2
Terminalbench Hard(Stanford × Laude Institute (2026))
0.1
Hle(Center for AI Safety + Scale AI (2025))
0.1

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input Price$1 / 1M tokens
Output Price$3 / 1M tokens
Blended Price (3:1)$1.5 / 1M tokens

Speed

Tokens/sec29.9
Time to First Token0.80s
Time to Answer67.67s

Provider Price Ranking

Provider Price Ranking

3 providers

Cheapest: Nous ResearchMost Expensive: Kilo Gateway
ProviderInputOutput
1Nous ResearchPRIMARY
$1
$3
2OpenRouter
$1
$1
3Kilo Gateway
$1
$1

Compare pricing across different API providers for this model.

External Sources