メインコンテンツへスキップ

Llama 3.1 Nemotron Ultra 253B v1 (Reasoning)

NVIDIALlamaオープンウエイトLlama 3.1 Community License

説明

A 253B parameter derivative of Meta Llama 3.1 405B Instruct, developed by NVIDIA using Neural Architecture Search (NAS) and vertical compression. It underwent multi-phase post-training (SFT for Math, Code, Reasoning, Chat, Tool Calling; RL with GRPO) to enhance reasoning and instruction-following. Optimized for accuracy/efficiency tradeoff on NVIDIA GPUs. Supports 128k context.

リリース日
2025-04-07
パラメータ
253.0B
コンテキスト長
128K
モダリティ
text

能力レーダー

30
general
64
coding
72
reasoning
53
science
70
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング386
37.0
AA
総合ランキング457
31.0
AA
科学355
42.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)89.5%自己申告

General

BFCL v274.1%自己申告

Math

MATH-50097.0%自己申告
AIME 202572.5%自己申告

Reasoning

GPQANYU + Cohere + Anthropic (2023)76.0%自己申告
LiveCodeBench66.3%自己申告

AA評価指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
95.2
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
82.5
Aime(MAA (Mathematical Association of America))
74.7
Gpqa(NYU + Cohere + Anthropic (2023))
72.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))
64.1
Math Index(Artificial Analysis)
63.7
Aime 25(MAA (Mathematical Association of America))
63.7
Ifbench(Google Research (2023))
38.2
Tau2(Sierra + U Toronto + Vector Institute (2025))
11.4
Intelligence Index(Artificial Analysis)
7.5
Hle(Center for AI Safety + Scale AI (2025))
7.4
Terminalbench Hard(Stanford × Laude Institute (2026))
2.3

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Chat
90
Instruction Following
90
Structured Output
90
Math
80
Physics
80
Reasoning
80
General
80
Biology
80
Chemistry
80
Code
70
Tool Calling
70

価格設定

入力価格無料
出力価格無料
混合価格(3:1)無料

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダーデータがありません

外部リンク