NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)
NVIDIA開源權重NVIDIA Open Model License Agreement · 商用許可
描述
Nemotron 3 Nano is a 31.6B hybrid MoE model optimized for fast, long‑context agentic reasoning. It mixes Mamba‑2 and Transformer layers with a sparse MoE router (~3.6B active params per token) to deliver up to 4× higher throughput than Nemotron 2 and strong accuracy across math, coding, and tools. It supports a 1M‑token context window, offers Reasoning ON/OFF and a thinking‑budget to control costs, and ships with open weights, data, and RL tooling (NeMo Gym/RL). Released Dec 15, 2025 under the NVIDIA Open Model License, it’s built as the efficient backbone for multi‑agent systems at scale.
發布日期
2025-12-15
參數規模
32.0B
上下文長度
131K
支援模態
text
能力雷達圖
22
general
33
coding
18
reasoning
27
science
50
agents
0
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Agents
Terminal-Bench
8.5%自報
Biology
GPQANYU + Cohere + Anthropic (2023)
75.0%自報
SciCode
33.3%自報
Code
SWE-Bench Verified
38.8%自報
Communication
Tau2 Retail
56.9%自報
Tau2 Airline
48.0%自報
Tau2 Telecom
42.2%自報
Multi-Challenge
38.5%自報
Creativity
Arena-Hard v2
67.7%自報
Finance
MMLU-Pro
78.3%自報
MMLU-ProX
59.5%自報
General
LiveCodeBench v6
68.3%自報
Language
WMT24++
86.2%自報
Math
AIME 2025
99.2%自報
Humanity's Last Exam
15.5%自報
AA 評測指數
(Artificial Analysis)Math Index(Artificial Analysis)13.3
Intelligence Index(Artificial Analysis)7.2
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.6
Gpqa(NYU + Cohere + Anthropic (2023))0.4
Ifbench(Google Research (2023))0.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.4
Tau2(Sierra + U Toronto + Vector Institute (2025))0.3
Scicode(UIUC + Argonne National Lab (2024))0.2
Aime 25(MAA (Mathematical Association of America))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Lcr(Artificial Analysis)0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
LLM Stats 分類評分
(LLM Stats (zeroeval))Legal70
Language70
Finance70
Healthcare70
Creativity70
Writing70
Math60
Physics50
Reasoning50
General50
Biology50
Chemistry50
Communication50
Tool Calling50
Frontend Development40
Code30
Vision20
Agents10
定價
輸入價格$0.05 / 1M tokens
輸出價格$0.2 / 1M tokens
混合價格(3:1)$0.088 / 1M tokens
速度
Tokens/秒202.9
首Token延遲0.39s
首回答延遲0.39s
供應商價格排行
供應商價格排行
8 個供應商
最便宜: DeepInfra最貴: NanoGPT
供應商輸入輸出
1DeepInfra最便宜
$0
$0
2NVIDIA主要
$0.05
$0.2
3OpenRouter
$0.05
$0.2
4Kilo Gateway
$0.05
$0.2
5Vercel AI Gateway
$0.05
$0.24
6Cortecs
$0.06
$0.24
7Venice AI
$0.075
$0.3
8NanoGPT
$0.17
$0.68
比較該模型在不同 API 供應商之間的定價。