Hermes 4 - Llama-3.1 405B (Reasoning)
Nous ResearchLlama
Release Date
2025-08-27
Parameters
—
Context Length
131K
Modalities
text
Capability Radar
31
general
69
coding
70
reasoning
54
science
70
agents
0
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 370 | 39.0 | AA |
| General Ranking | 439 | 32.0 | AA |
| Science | 316 | 45.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))No benchmark data available
AA Evaluation Indices
(Artificial Analysis)Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))82.9
Gpqa(NYU + Cohere + Anthropic (2023))72.7
Math Index(Artificial Analysis)69.7
Aime 25(MAA (Mathematical Association of America))69.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))68.6
Ifbench(Google Research (2023))32.7
Lcr(Artificial Analysis)22.3
Tau2(Sierra + U Toronto + Vector Institute (2025))22.2
Terminalbench Hard(Stanford × Laude Institute (2026))11.4
Hle(Center for AI Safety + Scale AI (2025))10.9
Intelligence Index(Artificial Analysis)7.5
LLM Stats Category Scores
(LLM Stats (zeroeval))No category score data available
Pricing
Input Price$1 / 1M tokens
Output Price$3 / 1M tokens
Blended Price (3:1)$1.5 / 1M tokens
Speed
Tokens/sec43.5
Time to First Token0.74s
Time to Answer46.68s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1Nous ResearchPRIMARY
$1
$3
Compare pricing across different API providers for this model.