Перейти к основному содержанию

NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)

NVIDIAОткрытые весаNVIDIA Open Model License Agreement · Коммерческое использование

Описание

Nemotron 3 Nano is a 31.6B hybrid MoE model optimized for fast, long‑context agentic reasoning. It mixes Mamba‑2 and Transformer layers with a sparse MoE router (~3.6B active params per token) to deliver up to 4× higher throughput than Nemotron 2 and strong accuracy across math, coding, and tools. It supports a 1M‑token context window, offers Reasoning ON/OFF and a thinking‑budget to control costs, and ships with open weights, data, and RL tooling (NeMo Gym/RL). Released Dec 15, 2025 under the NVIDIA Open Model License, it’s built as the efficient backbone for multi‑agent systems at scale.

Дата выхода
2025-12-15
Параметры
32.0B
Длина контекста
131K
Модальности
text

Радар способностей

22
general
36
coding
18
reasoning
29
science
50
agents
0
multimodal

Рейтинги

Домен#МестоОценкаИсточник
Рейтинг кодинга489
23.0
AA
Общий рейтинг503
28.0
AA
Наука564
21.0
AA

Оценки бенчмарков (LLM Stats)

(LLM Stats (zeroeval))

Chat

Tau2 Airline48.0%Сам.
Multi-Challenge38.5%Сам.

Communication

Tau2 Retail56.9%Сам.
Tau2 Telecom42.2%Сам.

General

Arena-Hard v267.7%Сам.

Language

WMT24++86.2%Сам.
MMLU-Pro78.3%Сам.
MMLU-ProX59.5%Сам.

Math

AIME 202599.2%Сам.

Reasoning

GPQANYU + Cohere + Anthropic (2023)75.0%Сам.
LiveCodeBench v668.3%Сам.
SWE-Bench Verified38.8%Сам.
SciCode33.3%Сам.
Humanity's Last Exam15.5%Сам.
Terminal-Bench8.5%Сам.

Индексы оценки AA

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
57.9
Gpqa(NYU + Cohere + Anthropic (2023))
39.9
Ifbench(Google Research (2023))
37.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))
36.0
Tau2(Sierra + U Toronto + Vector Institute (2025))
25.4
Aime 25(MAA (Mathematical Association of America))
13.3
Math Index(Artificial Analysis)
13.3
Terminalbench Hard(Stanford × Laude Institute (2026))
12.1
Lcr(Artificial Analysis)
10.7
Intelligence Index(Artificial Analysis)
6.8
Hle(Center for AI Safety + Scale AI (2025))
4.6

Оценки категорий LLM Stats

(LLM Stats (zeroeval))
Language
70
Legal
70
Finance
70
Healthcare
70
Creativity
70
Writing
70
Math
60
Physics
50
Reasoning
50
General
50
Biology
50
Chemistry
50
Communication
50
Tool Calling
50
Chat
40
Frontend Development
40
Code
30
Vision
20
Agents
10

Цены

Цена ввода$0.05 / 1M токенов
Цена вывода$0.2 / 1M токенов
Смешанная цена (3:1)$0.088 / 1M токенов

Скорость

Токенов/сек220.8
Задержка первого токена0.38s
Время до первого ответа0.38s

Рейтинг цен провайдеров

Рейтинг цен провайдеров

8 провайдеров

Самый дешевый: DeepInfraСамый дорогой: NanoGPT
ПровайдерВводВывод
1DeepInfraСамый дешевый
$0
$0
2NVIDIAОсновной
$0.05
$0.2
3OpenRouter
$0.05
$0.2
4Kilo Gateway
$0.05
$0.2
5Vercel AI Gateway
$0.05
$0.2
6Cortecs
$0.06
$0.24
7Venice AI
$0.075
$0.3
8NanoGPT
$0.17
$0.68

Сравнение цен разных API-провайдеров для этой модели.

Внешние ссылки