Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)
NVIDIALlamaOpen WeightLlama 3.1 Community License
Description
Llama-3.1-Nemotron-Nano-8B-v1 is a large language model (LLM) which is a derivative of Meta Llama-3.1-8B-Instruct (AKA the reference model). It is a reasoning model that is post trained for reasoning, human chat preferences, and tasks, such as RAG and tool calling.
Release Date
2025-05-20
Parameters
8.0B
Context Length
131K
Modalities
text
Capability Radar
22
general
41
coding
61
reasoning
23
science
60
agents
0
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 374 | 27.0 | AA |
| General Ranking | 489 | 23.0 | AA |
| Science | 507 | 20.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
54.1%SR
Communication
MT-Bench
0.81 / 100SR
General
MBPP
0.85 / 100SR
IFEvalGoogle Research (2023)
79.3%SR
BFCL v2
63.6%SR
Math
MATH-500
95.4%SR
AIME 2025
47.1%SR
AA Evaluation Indices
(Artificial Analysis)Math Index(Artificial Analysis)50.0
Intelligence Index(Artificial Analysis)8.4
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.9
Aime(MAA (Mathematical Association of America))0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.6
Aime 25(MAA (Mathematical Association of America))0.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.5
Gpqa(NYU + Cohere + Anthropic (2023))0.4
Ifbench(Google Research (2023))0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))0.1
Scicode(UIUC + Argonne National Lab (2024))0.1
Hle(Center for AI Safety + Scale AI (2025))0.1
Lcr(Artificial Analysis)0.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Roleplay80
Structured Output80
Instruction Following80
Communication80
Creativity80
Math70
Reasoning70
General70
Tool Calling60
Physics50
Biology50
Chemistry50
Pricing
Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
No provider data available