Skip to main content

Hermes 4 - Llama-3.1 405B (Non-reasoning)

Nous ResearchLlama
Release Date
2025-08-27
Parameters
—
Context Length
131K
Modalities
text

Capability Radar

27
general
55
coding
22
reasoning
38
science
32
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking412
33.0
AA
General Ranking456
31.0
AA
Science488
28.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
72.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))
54.6
Gpqa(NYU + Cohere + Anthropic (2023))
53.6
Ifbench(Google Research (2023))
34.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
26.6
Lcr(Artificial Analysis)
22.0
Aime 25(MAA (Mathematical Association of America))
15.3
Math Index(Artificial Analysis)
15.3
Terminalbench Hard(Stanford × Laude Institute (2026))
9.8
Intelligence Index(Artificial Analysis)
7.4
Hle(Center for AI Safety + Scale AI (2025))
4.2

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input Price$1 / 1M tokens
Output Price$3 / 1M tokens
Blended Price (3:1)$1.5 / 1M tokens

Speed

Tokens/sec42.3
Time to First Token0.71s
Time to Answer0.71s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1Nous ResearchPRIMARY
$1
$3

Compare pricing across different API providers for this model.

External Sources