o1-preview
OpenAIOpenAI o-seriesProprietary
Description
A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.
Release Date
2024-09-12
Parameters
—
Context Length
200K
Modalities
image, pdf, text
Capability Radar
37
general
34
coding
82
reasoning
77
science
68
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 355 | 42.0 | AA |
| General Ranking | 330 | 41.0 | AA |
| Science | 83 | 77.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Factuality
SimpleQA
42.4%SR
General
MMLU
90.8%SR
Math
MGSM
90.8%SR
MATH
85.5%SR
LiveBench
52.3%SR
AIME 2024
42.0%SR
Reasoning
GPQANYU + Cohere + Anthropic (2023)
73.3%SR
SWE-Bench Verified
41.3%SR
AA Evaluation Indices
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))92.4
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))84.8
Math Index(Artificial Analysis)79.7
Aime 25(MAA (Mathematical Association of America))79.7
Gpqa(NYU + Cohere + Anthropic (2023))76.5
Coding Index(Artificial Analysis)34.0
Intelligence Index(Artificial Analysis)11.4
LLM Stats Category Scores
(LLM Stats (zeroeval))Language90
Legal90
Finance90
Healthcare90
Math70
Physics70
Biology70
Chemistry70
Reasoning60
General60
Factuality40
Frontend Development40
Code40
Pricing
Input Price$16.5 / 1M tokens
Output Price$66 / 1M tokens
Blended Price (3:1)$28.875 / 1M tokens
Cache Read Price$7.5 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1OpenAIPRIMARY
$16.5
$66
Compare pricing across different API providers for this model.