NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)
NVIDIAオープンウエイトNVIDIA Open Model License Agreement · 商用利用可
説明
Nemotron 3 Nano is a 31.6B hybrid MoE model optimized for fast, long‑context agentic reasoning. It mixes Mamba‑2 and Transformer layers with a sparse MoE router (~3.6B active params per token) to deliver up to 4× higher throughput than Nemotron 2 and strong accuracy across math, coding, and tools. It supports a 1M‑token context window, offers Reasoning ON/OFF and a thinking‑budget to control costs, and ships with open weights, data, and RL tooling (NeMo Gym/RL). Released Dec 15, 2025 under the NVIDIA Open Model License, it’s built as the efficient backbone for multi‑agent systems at scale.
リリース日
2025-12-15
パラメータ
32.0B
コンテキスト長
131K
モダリティ
text
能力レーダー
22
general
33
coding
18
reasoning
27
science
50
agents
0
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 114 | 36.0 | LS |
| コーディングランキング | 408 | 22.0 | AA |
| 総合ランキング | 436 | 29.0 | AA |
| 科学 | 466 | 26.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
Terminal-Bench
8.5%自己申告
Biology
GPQANYU + Cohere + Anthropic (2023)
75.0%自己申告
SciCode
33.3%自己申告
Code
SWE-Bench Verified
38.8%自己申告
Communication
Tau2 Retail
56.9%自己申告
Tau2 Airline
48.0%自己申告
Tau2 Telecom
42.2%自己申告
Multi-Challenge
38.5%自己申告
Creativity
Arena-Hard v2
67.7%自己申告
Finance
MMLU-Pro
78.3%自己申告
MMLU-ProX
59.5%自己申告
General
LiveCodeBench v6
68.3%自己申告
Language
WMT24++
86.2%自己申告
Math
AIME 2025
99.2%自己申告
Humanity's Last Exam
15.5%自己申告
AA評価指数
(Artificial Analysis)Math Index(Artificial Analysis)13.3
Intelligence Index(Artificial Analysis)7.2
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.6
Gpqa(NYU + Cohere + Anthropic (2023))0.4
Ifbench(Google Research (2023))0.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.4
Tau2(Sierra + U Toronto + Vector Institute (2025))0.3
Scicode(UIUC + Argonne National Lab (2024))0.2
Aime 25(MAA (Mathematical Association of America))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Lcr(Artificial Analysis)0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Legal70
Language70
Finance70
Healthcare70
Creativity70
Writing70
Math60
Physics50
Reasoning50
General50
Biology50
Chemistry50
Communication50
Tool Calling50
Frontend Development40
Code30
Vision20
Agents10
価格設定
入力価格$0.05 / 1Mトークン
出力価格$0.2 / 1Mトークン
混合価格(3:1)$0.088 / 1Mトークン
速度
トークン/秒202.9
初トークン遅延0.39s
初回答遅延0.39s
プロバイダー価格ランキング
プロバイダー価格ランキング
8 プロバイダー
最安: DeepInfra最高: NanoGPT
プロバイダー入力出力
1DeepInfra最安
$0
$0
2NVIDIAプライマリ
$0.05
$0.2
3OpenRouter
$0.05
$0.2
4Kilo Gateway
$0.05
$0.2
5Vercel AI Gateway
$0.05
$0.24
6Cortecs
$0.06
$0.24
7Venice AI
$0.075
$0.3
8NanoGPT
$0.17
$0.68
このモデルの異なるAPIプロバイダー間の価格を比較。