Skip to main content

Nova Micro

AmazonAmazonProprietary

Description

A text-only model that delivers lowest-latency responses at very low cost while maintaining strong performance on core language tasks. Optimized for speed and efficiency while preserving high accuracy on key benchmarks.

Release Date
2024-12-03
Parameters
Context Length
128K
Modalities
text

Capability Radar

19
general
13
coding
19
reasoning
20
science
60
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking503
10.0
AA
General Ranking497
21.0
AA
Science507
18.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)40.0%SR

Code

HumanEvalOpenAI (2021)81.1%SR

Economics

FinQA65.2%SR
CRAG43.1%SR

Finance

MMLU77.6%SR

General

ARC-C90.2%SR
IFEvalGoogle Research (2023)87.2%SR
BFCL56.2%SR

Language

Translation Set1→en COMET2288.7%SR
Translation en→Set1 COMET2288.5%SR
BBH79.5%SR
Translation Set1→en spBleu42.6%SR
Translation en→Set1 spBleu40.2%SR
SQuALITY18.8%SR

Math

GSM8k92.3%SR
DROP79.3%SR
MATH69.3%SR

AA Evaluation Indices

(Artificial Analysis)
Math Index(Artificial Analysis)
6.0
Intelligence Index(Artificial Analysis)
4.4
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.5
Gpqa(NYU + Cohere + Anthropic (2023))
0.4
Ifbench(Google Research (2023))
0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.1
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.1
Lcr(Artificial Analysis)
0.1
Scicode(UIUC + Argonne National Lab (2024))
0.1
Aime(MAA (Mathematical Association of America))
0.1
Aime 25(MAA (Mathematical Association of America))
0.1
Hle(Center for AI Safety + Scale AI (2025))
0.0
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Instruction Following
90
Structured Output
90
Legal
80
Math
80
Healthcare
80
Code
80
Reasoning
70
General
70
Language
60
Finance
60
Tool Calling
60
Economics
50
Physics
40
Search
40
Biology
40
Chemistry
40
Long Context
20
Summarization
20

Pricing

Input Price$0.035 / 1M tokens
Output Price$0.14 / 1M tokens
Blended Price (3:1)$0.061 / 1M tokens
Cache Read Price$0.00875 / 1M tokens

Speed

Tokens/sec286.7
Time to First Token0.69s
Time to Answer0.69s

Provider Price Ranking

Provider Price Ranking

5 providers

Cheapest: AmazonMost Expensive: NanoGPT
ProviderInputOutput
1AmazonPRIMARY
$0.035
$0.14
2OpenRouter
$0.035
$0.14
3Kilo Gateway
$0.035
$0.14
4Amazon Bedrock
$0.035
$0.14
5NanoGPT
$0.0357
$0.1394

Compare pricing across different API providers for this model.

External Sources