Skip to main content

o4-mini (high)

OpenAIOpenAI o-seriesProprietary

Description

o4-mini is OpenAI's latest small o-series model, optimized for fast, effective reasoning with exceptionally efficient performance in coding and visual tasks. It is faster and more affordable than o3.

Release Date
2025-04-16
Parameters
—
Context Length
200K
Modalities
image, text

Capability Radar

37
general
86
coding
92
reasoning
59
science
60
agents
85
multimodal

Rankings

Domain#RankScoreSource
Code Ranking227
62.0
AA
General Ranking193
55.0
AA
Multimodal Ranking38
60.0
LS
Science240
53.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Chat

TAU-bench Retail71.8%SR
Multi-Challenge43.0%SR

General

Aider-Polyglot68.9%SR
Aider-Polyglot Edit58.2%SR

Math

AIME 202493.4%SR
AIME 202592.7%SR
MathVista84.3%SR

Multimodal

MMMU81.6%SR

Reasoning

GPQANYU + Cohere + Anthropic (2023)81.4%SR
CharXiv-R72.0%SR
SWE-Bench Verified68.1%SR
BrowseCompOpenAI (2025)51.5%SR
TAU-bench Airline49.2%SR
Humanity's Last Exam14.7%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
98.9
Aime(MAA (Mathematical Association of America))
94.0
Math Index(Artificial Analysis)
90.7
Aime 25(MAA (Mathematical Association of America))
90.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
85.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
83.2
Gpqa(NYU + Cohere + Anthropic (2023))
78.4
Ifbench(Google Research (2023))
68.7
Lcr(Artificial Analysis)
61.0
Tau2(Sierra + U Toronto + Vector Institute (2025))
55.6
Intelligence Index(Artificial Analysis)
16.7
Hle(Center for AI Safety + Scale AI (2025))
16.5
Terminalbench Hard(Stanford × Laude Institute (2026))
15.2

LLM Stats Category Scores

(LLM Stats (zeroeval))
Multimodal
80
Physics
80
Healthcare
80
Biology
80
Chemistry
80
Math
70
Reasoning
70
Frontend Development
70
General
70
Code
70
Chat
60
Tool Calling
60
Vision
60
Search
50
Agents
50
Communication
50

Pricing

Input Price$1.1 / 1M tokens
Output Price$4.4 / 1M tokens
Blended Price (3:1)$1.925 / 1M tokens
Cache Read Price$0.275 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

19 providers

Cheapest: PoeMost Expensive: LLM Gateway
ProviderInputOutput
1PoeCheapest
$0.99
$4
2OpenAIPRIMARY
$1.1
$4.4
3NanoGPT
$1.1
$4.4
4Abacus
$1.1
$4.4
5OpenRouter
$1.1
$4.4
6Jiekou.AI
$1.1
$4.4
7Kilo Gateway
$1.1
$4.4
8Cloudflare AI Gateway
$1.1
$4.4
9Helicone
$1.1
$4.4
10AIHubMix
$1.1
$4.4
11Azure Cognitive Services
$1.1
$4.4
12Vercel AI Gateway
$1.1
$4.4
13DevPass (LLM Gateway)
$1.1
$4.4
14Azure
$1.1
$4.4
15NEAR AI Cloud
$1.1
$4.4
16Merge Gateway
$1.1
$4.4
17Impossibl
$1.1
$4.4
18Eden AI
$1.1
$4.4
19LLM Gateway
$1.1
$4.4

Compare pricing across different API providers for this model.

External Sources