Skip to main content

Nova Micro

AmazonAmazonProprietary

Description

A text-only model that delivers lowest-latency responses at very low cost while maintaining strong performance on core language tasks. Optimized for speed and efficiency while preserving high accuracy on key benchmarks.

Release Date
2024-12-03
Parameters
—
Context Length
128K
Modalities
text

Capability Radar

20
general
14
coding
19
reasoning
26
science
60
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking578
11.0
AA
General Ranking570
21.0
AA
Science586
18.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)87.2%SR

Finance

FinQA65.2%SR

General

MMLU77.6%SR
BFCL56.2%SR

Language

Translation Set1→en COMET2288.7%SR
Translation en→Set1 COMET2288.5%SR
Translation Set1→en spBleu42.6%SR
Translation en→Set1 spBleu40.2%SR

Math

GSM8k92.3%SR
MATH69.3%SR

Reasoning

ARC-C90.2%SR
HumanEvalOpenAI (2021)81.1%SR
BBH79.5%SR
DROP79.3%SR
CRAG43.1%SR
GPQANYU + Cohere + Anthropic (2023)40.0%SR

Summarization

SQuALITY18.8%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
70.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
53.1
Gpqa(NYU + Cohere + Anthropic (2023))
35.8
Ifbench(Google Research (2023))
29.4
Tau2(Sierra + U Toronto + Vector Institute (2025))
14.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))
14.0
Lcr(Artificial Analysis)
13.0
Aime(MAA (Mathematical Association of America))
8.0
Math Index(Artificial Analysis)
6.0
Aime 25(MAA (Mathematical Association of America))
6.0
Intelligence Index(Artificial Analysis)
5.9
Hle(Center for AI Safety + Scale AI (2025))
4.6
Terminalbench Hard(Stanford × Laude Institute (2026))
1.5

LLM Stats Category Scores

(LLM Stats (zeroeval))
Chat
90
Instruction Following
90
Structured Output
90
Legal
80
Math
80
Healthcare
80
Code
80
Reasoning
70
General
70
Language
60
Finance
60
Tool Calling
60
Economics
50
Physics
40
Search
40
Biology
40
Chemistry
40
Long Context
20
Summarization
20

Pricing

Input Price$0.035 / 1M tokens
Output Price$0.14 / 1M tokens
Blended Price (3:1)$0.061 / 1M tokens
Cache Read Price$0.00925 / 1M tokens
Cache Write Price$0.037 / 1M tokens

Speed

Tokens/sec239.5
Time to First Token0.58s
Time to Answer0.58s

Provider Price Ranking

Provider Price Ranking

4 providers

Cheapest: AmazonMost Expensive: Amazon Bedrock
ProviderInputOutput
1AmazonPRIMARY
$0.035
$0.14
2OpenRouter
$0.035
$0.14
3Kilo Gateway
$0.035
$0.14
4Amazon Bedrock
$0.037
$0.148

Compare pricing across different API providers for this model.

External Sources