o4-mini (high)
OpenAIOpenAI o-seriesProprietary
Description
o4-mini is OpenAI's latest small o-series model, optimized for fast, effective reasoning with exceptionally efficient performance in coding and visual tasks. It is faster and more affordable than o3.
Release Date
2025-04-16
Parameters
—
Context Length
200K
Modalities
image, text
Capability Radar
42
general
77
coding
92
reasoning
55
science
60
agents
85
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 82 | 44.0 | LS |
| Code Ranking | 153 | 63.0 | AA |
| General Ranking | 158 | 61.0 | AA |
| Multimodal Ranking | 53 | 45.0 | LS |
| Science | 154 | 62.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
BrowseCompOpenAI (2025)
51.5%SR
Biology
GPQANYU + Cohere + Anthropic (2023)
81.4%SR
Code
Aider-Polyglot
68.9%SR
SWE-Bench Verified
68.1%SR
Aider-Polyglot Edit
58.2%SR
Communication
TAU-bench Retail
71.8%SR
TAU-bench Airline
49.2%SR
Multi-Challenge
43.0%SR
General
MMMU
81.6%SR
Math
AIME 2024
93.4%SR
AIME 2025
92.7%SR
MathVista
84.3%SR
Humanity's Last Exam
14.7%SR
Multimodal
CharXiv-R
72.0%SR
AA Evaluation Indices
(Artificial Analysis)Math Index(Artificial Analysis)90.7
Intelligence Index(Artificial Analysis)26.1
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))1.0
Aime(MAA (Mathematical Association of America))0.9
Aime 25(MAA (Mathematical Association of America))0.9
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.8
Ifbench(Google Research (2023))0.7
Lcr(Artificial Analysis)0.6
Tau2(Sierra + U Toronto + Vector Institute (2025))0.6
Scicode(UIUC + Argonne National Lab (2024))0.5
Hle(Center for AI Safety + Scale AI (2025))0.2
Terminalbench Hard(Stanford × Laude Institute (2026))0.2
LLM Stats Category Scores
(LLM Stats (zeroeval))Multimodal80
Physics80
Healthcare80
Biology80
Chemistry80
Math70
Reasoning70
Frontend Development70
General70
Code70
Tool Calling60
Vision60
Search50
Agents50
Communication50
Pricing
Input Price$1.1 / 1M tokens
Output Price$4.4 / 1M tokens
Blended Price (3:1)$1.925 / 1M tokens
Cache Read Price$0.275 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
16 providers
Cheapest: PoeMost Expensive: Impossibl
ProviderInputOutput
1PoeCheapest
$0.99
$4
2OpenAIPRIMARY
$1.1
$4.4
3NanoGPT
$1.1
$4.4
4Abacus
$1.1
$4.4
5OpenRouter
$1.1
$4.4
6Jiekou.AI
$1.1
$4.4
7Kilo Gateway
$1.1
$4.4
8Cloudflare AI Gateway
$1.1
$4.4
9Helicone
$1.1
$4.4
10Azure Cognitive Services
$1.1
$4.4
11Vercel AI Gateway
$1.1
$4.4
12LLM Gateway
$1.1
$4.4
13Azure
$1.1
$4.4
14NEAR AI Cloud
$1.1
$4.4
15Merge Gateway
$1.1
$4.4
16Impossibl
$1.1
$4.4
Compare pricing across different API providers for this model.