Skip to main content

Llama 3.3 Nemotron Super 49B v1 (Non-reasoning)

NVIDIALlamaOpen WeightLlama 3.1 Community License

Description

Llama-3.3-Nemotron-Super-49B-v1 is a large language model (LLM) derived from Meta Llama-3.3-70B-Instruct. It's post-trained for reasoning, chat, RAG, and tool calling, offering a balance between accuracy and efficiency (optimized for single H100). It underwent multi-phase post-training including SFT and RL (RLOO, RPO).

Release Date
2025-03-18
Parameters
49.9B
Context Length
131K
Modalities
text

Capability Radar

26
general
28
coding
25
reasoning
37
science
70
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking540
15.0
AA
General Ranking420
33.0
AA
Science502
26.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Chat

MT-Bench0.92 / 100SR

General

Arena Hard88.3%SR
BFCL v273.7%SR

Math

MATH-50096.6%SR
AIME 202558.4%SR

Reasoning

MBPP0.91 / 100SR
GPQANYU + Cohere + Anthropic (2023)66.7%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
77.5
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
69.8
Gpqa(NYU + Cohere + Anthropic (2023))
51.7
Ifbench(Google Research (2023))
39.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))
28.0
Aime(MAA (Mathematical Association of America))
19.3
Lcr(Artificial Analysis)
12.0
Math Index(Artificial Analysis)
7.7
Aime 25(MAA (Mathematical Association of America))
7.7
Intelligence Index(Artificial Analysis)
7.3
Hle(Center for AI Safety + Scale AI (2025))
3.8
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Chat
90
Roleplay
90
Communication
90
Creativity
90
Writing
90
Math
80
Reasoning
80
General
80
Physics
70
Biology
70
Chemistry
70
Tool Calling
70

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

No provider data available

External Sources