メインコンテンツへスキップ

Nemotron 3 Ultra 550B A55B (Reasoning)

NVIDIAオープンウエイトOpenMDW License v1.1 · 商用利用可

説明

Nemotron 3 Ultra is NVIDIA's frontier-scale open model with 550B total / 55B active parameters, built for agentic reasoning, long-context analysis, tool use, and high-stakes RAG. It uses a hybrid Latent Mixture-of-Experts (LatentMoE) architecture interleaving Mamba-2, MoE, and select Attention layers, with Multi-Token Prediction (MTP) for native speculative decoding, and is pre-trained on ~20T tokens with an NVFP4 recipe. Reasoning is configurable on/off (plus a medium-effort mode) via the chat template. It supports up to a 1M-token context and 10 languages (English, French, Spanish, Italian, German, Japanese, Hindi, Korean, Brazilian Portuguese, Chinese). Released with open weights, training data, and recipes under the OpenMDW-1.1 license.

リリース日
2026-06-04
パラメータ
550.0B
コンテキスト長
1.0M
モダリティ
text

能力レーダー

36
general
48
coding
87
reasoning
59
science
40
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
エージェント能力92
26.0
LS
コーディングランキング146
67.0
AA
総合ランキング82
75.0
AA
科学115
69.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

Finance Agent53.7%自己申告
Finance Agent v237.5%
TAU3-Bench22.6%自己申告

Chat

Multi-Challenge63.8%自己申告

Code

PinchBench90.0%自己申告

General

GDPval46.7%自己申告

Instruction Following

IFBench81.7%自己申告

Language

MMLU-Pro86.8%自己申告
WMT24++83.7%自己申告
MMLU-ProX83.0%自己申告

Long Context

RULER94.7%自己申告
LongBench v261.9%自己申告

Math

IMO-AnswerBench92.3%自己申告

Reasoning

LiveCodeBench v689.0%自己申告
GPQANYU + Cohere + Anthropic (2023)87.0%自己申告
Apex84.8%自己申告
SWE-Bench Verified70.7%自己申告
SWE-bench Multilingual67.7%自己申告
AA-LCR65.4%自己申告
Terminal-Bench 2.156.4%自己申告
ProfBench56.0%自己申告
SciCode44.6%自己申告
BrowseCompOpenAI (2025)44.4%自己申告
Humanity's Last Exam37.4%自己申告
CritPT3.1%自己申告

Science

OmniScience78.7%自己申告

AA評価指数

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
86.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
83.3
Ifbench(Google Research (2023))
81.4
Lcr(Artificial Analysis)
71.0
Terminalbench V2 1
53.9
Coding Index(Artificial Analysis)
49.3
Scicode(UIUC + Argonne National Lab (2024))
39.9
Intelligence Index(Artificial Analysis)
38.3
Terminalbench Hard(Stanford × Laude Institute (2026))
36.4
Hle(Center for AI Safety + Scale AI (2025))
28.4
Tau Banking
14.2

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Instruction Following
80
Language
80
Science
80
Healthcare
80
Knowledge
70
Legal
70
Long Context
70
Physics
70
Frontend Development
70
Biology
70
Chemistry
70
Code
70
Chat
60
Math
60
Reasoning
60
Structured Output
60
Finance
60
General
60
Communication
60
Agents
50
Search
40
Tool Calling
40
Vision
40

価格設定

入力価格$0.675 / 1Mトークン
出力価格$2.675 / 1Mトークン
混合価格(3:1)$1.175 / 1Mトークン
キャッシュ読み取り価格$0.15 / 1Mトークン

速度

トークン/秒87.1
初トークン遅延2.23s
初回答遅延28.36s

プロバイダー価格ランキング

プロバイダー価格ランキング

7 プロバイダー

最安: Nvidia最高: Venice AI
プロバイダー入力出力
1Nvidia最安
$0.5
$2.5
2NanoGPT
$0.5
$2.5
3OpenRouter
$0.5
$2.2
4Kilo Gateway
$0.5
$2.2
5Vercel AI Gateway
$0.6
$2.4
6Together AI
$0.6
$3.6
7Venice AI
$0.625
$3.125

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク