Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)
NVIDIALlamaOpen WeightLlama 3.1 Community License
Description
Llama-3.1-Nemotron-Nano-8B-v1 is a large language model (LLM) which is a derivative of Meta Llama-3.1-8B-Instruct (AKA the reference model). It is a reasoning model that is post trained for reasoning, human chat preferences, and tasks, such as RAG and tool calling.
Release Date
2025-05-20
Parameters
8.0B
Context Length
131K
Modalities
text
Capability Radar
21
general
49
coding
61
reasoning
30
science
60
agents
0
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 447 | 27.0 | AA |
| General Ranking | 562 | 21.0 | AA |
| Science | 542 | 22.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Chat
MT-Bench
0.81 / 100SR
IFEvalGoogle Research (2023)
79.3%SR
General
BFCL v2
63.6%SR
Math
MATH-500
95.4%SR
AIME 2025
47.1%SR
Reasoning
MBPP
0.85 / 100SR
GPQANYU + Cohere + Anthropic (2023)
54.1%SR
AA Evaluation Indices
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))94.7
Aime(MAA (Mathematical Association of America))70.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))55.6
Math Index(Artificial Analysis)50.0
Aime 25(MAA (Mathematical Association of America))50.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))49.3
Gpqa(NYU + Cohere + Anthropic (2023))40.8
Ifbench(Google Research (2023))25.5
Tau2(Sierra + U Toronto + Vector Institute (2025))11.7
Intelligence Index(Artificial Analysis)7.3
Hle(Center for AI Safety + Scale AI (2025))5.2
Lcr(Artificial Analysis)0.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Chat80
Instruction Following80
Roleplay80
Structured Output80
Communication80
Creativity80
Math70
Reasoning70
General70
Tool Calling60
Physics50
Biology50
Chemistry50
Pricing
Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
No provider data available