o1-preview
OpenAIOpenAI o-seriesProprietary
Description
A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.
Release Date
2024-09-12
Parameters
—
Context Length
200K
Modalities
image, pdf, text
Capability Radar
41
general
34
coding
82
reasoning
77
science
68
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 276 | 43.0 | AA |
| General Ranking | 247 | 49.0 | AA |
| Science | 64 | 78.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
73.3%SR
Code
SWE-Bench Verified
41.3%SR
Factuality
SimpleQA
42.4%SR
Finance
MMLU
90.8%SR
General
LiveBench
52.3%SR
Math
MGSM
90.8%SR
MATH
85.5%SR
AIME 2024
42.0%SR
AA Evaluation Indices
(Artificial Analysis)Math Index(Artificial Analysis)79.3
Coding Index(Artificial Analysis)34.0
Intelligence Index(Artificial Analysis)17.2
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Aime 25(MAA (Mathematical Association of America))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.8
LLM Stats Category Scores
(LLM Stats (zeroeval))Legal90
Language90
Finance90
Healthcare90
Math70
Physics70
Biology70
Chemistry70
Reasoning60
General60
Factuality40
Frontend Development40
Code40
Pricing
Input Price$16.5 / 1M tokens
Output Price$66 / 1M tokens
Blended Price (3:1)$28.875 / 1M tokens
Cache Read Price$7.5 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1OpenAIPRIMARY
$16.5
$66
Compare pricing across different API providers for this model.