Llama 3.1 Nemotron Ultra 253B v1 (Reasoning)
NVIDIALlama오픈 웨이트Llama 3.1 Community License
설명
A 253B parameter derivative of Meta Llama 3.1 405B Instruct, developed by NVIDIA using Neural Architecture Search (NAS) and vertical compression. It underwent multi-phase post-training (SFT for Math, Code, Reasoning, Chat, Tool Calling; RL with GRPO) to enhance reasoning and instruction-following. Optimized for accuracy/efficiency tradeoff on NVIDIA GPUs. Supports 128k context.
출시일
2025-04-07
파라미터
253.0B
컨텍스트 길이
128K
모달리티
text
능력 레이더
30
general
64
coding
72
reasoning
53
science
70
agents
0
multimodal
랭킹
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
89.5%자체 보고
General
BFCL v2
74.1%자체 보고
Math
MATH-500
97.0%자체 보고
AIME 2025
72.5%자체 보고
Reasoning
GPQANYU + Cohere + Anthropic (2023)
76.0%자체 보고
LiveCodeBench
66.3%자체 보고
AA 평가 지수
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))95.2
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))82.5
Aime(MAA (Mathematical Association of America))74.7
Gpqa(NYU + Cohere + Anthropic (2023))72.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))64.1
Math Index(Artificial Analysis)63.7
Aime 25(MAA (Mathematical Association of America))63.7
Ifbench(Google Research (2023))38.2
Tau2(Sierra + U Toronto + Vector Institute (2025))11.4
Intelligence Index(Artificial Analysis)7.5
Hle(Center for AI Safety + Scale AI (2025))7.4
Terminalbench Hard(Stanford × Laude Institute (2026))2.3
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Chat90
Instruction Following90
Structured Output90
Math80
Physics80
Reasoning80
General80
Biology80
Chemistry80
Code70
Tool Calling70
가격
입력 가격무료
출력 가격무료
혼합 가격 (3:1)무료
속도
토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s
공급자 가격 순위
프로바이더 데이터가 없습니다