Hermes 4 - Llama-3.1 405B (Non-reasoning)
Nous ResearchLlama
Release Date
2025-08-27
Parameters
—
Context Length
131K
Modalities
text
Capability Radar
27
general
55
coding
22
reasoning
38
science
32
agents
0
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 412 | 33.0 | AA |
| General Ranking | 456 | 31.0 | AA |
| Science | 488 | 28.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))No benchmark data available
AA Evaluation Indices
(Artificial Analysis)Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))72.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))54.6
Gpqa(NYU + Cohere + Anthropic (2023))53.6
Ifbench(Google Research (2023))34.8
Tau2(Sierra + U Toronto + Vector Institute (2025))26.6
Lcr(Artificial Analysis)22.0
Aime 25(MAA (Mathematical Association of America))15.3
Math Index(Artificial Analysis)15.3
Terminalbench Hard(Stanford × Laude Institute (2026))9.8
Intelligence Index(Artificial Analysis)7.4
Hle(Center for AI Safety + Scale AI (2025))4.2
LLM Stats Category Scores
(LLM Stats (zeroeval))No category score data available
Pricing
Input Price$1 / 1M tokens
Output Price$3 / 1M tokens
Blended Price (3:1)$1.5 / 1M tokens
Speed
Tokens/sec42.3
Time to First Token0.71s
Time to Answer0.71s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1Nous ResearchPRIMARY
$1
$3
Compare pricing across different API providers for this model.