Skip to main content

o1-preview

OpenAIOpenAI o-seriesProprietary

Description

A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.

Release Date
2024-09-12
Parameters
Context Length
200K
Modalities
image, pdf, text

Capability Radar

41
general
34
coding
82
reasoning
77
science
68
agents
80
multimodal

Rankings

Domain#RankScoreSource
Code Ranking276
43.0
AA
General Ranking247
49.0
AA
Science64
78.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)73.3%SR

Code

SWE-Bench Verified41.3%SR

Factuality

SimpleQA42.4%SR

Finance

MMLU90.8%SR

General

LiveBench52.3%SR

Math

MGSM90.8%SR
MATH85.5%SR
AIME 202442.0%SR

AA Evaluation Indices

(Artificial Analysis)
Math Index(Artificial Analysis)
79.3
Coding Index(Artificial Analysis)
34.0
Intelligence Index(Artificial Analysis)
17.2
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Aime 25(MAA (Mathematical Association of America))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.8

LLM Stats Category Scores

(LLM Stats (zeroeval))
Legal
90
Language
90
Finance
90
Healthcare
90
Math
70
Physics
70
Biology
70
Chemistry
70
Reasoning
60
General
60
Factuality
40
Frontend Development
40
Code
40

Pricing

Input Price$16.5 / 1M tokens
Output Price$66 / 1M tokens
Blended Price (3:1)$28.875 / 1M tokens
Cache Read Price$7.5 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1OpenAIPRIMARY
$16.5
$66

Compare pricing across different API providers for this model.

External Sources