Skip to main content

o1-preview

OpenAIOpenAI o-seriesProprietary

Description

A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.

Release Date
2024-09-12
Parameters
—
Context Length
200K
Modalities
image, pdf, text

Capability Radar

37
general
34
coding
82
reasoning
77
science
68
agents
80
multimodal

Rankings

Domain#RankScoreSource
Code Ranking355
42.0
AA
General Ranking330
41.0
AA
Science83
77.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Factuality

SimpleQA42.4%SR

General

MMLU90.8%SR

Math

MGSM90.8%SR
MATH85.5%SR
LiveBench52.3%SR
AIME 202442.0%SR

Reasoning

GPQANYU + Cohere + Anthropic (2023)73.3%SR
SWE-Bench Verified41.3%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
92.4
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
84.8
Math Index(Artificial Analysis)
79.7
Aime 25(MAA (Mathematical Association of America))
79.7
Gpqa(NYU + Cohere + Anthropic (2023))
76.5
Coding Index(Artificial Analysis)
34.0
Intelligence Index(Artificial Analysis)
11.4

LLM Stats Category Scores

(LLM Stats (zeroeval))
Language
90
Legal
90
Finance
90
Healthcare
90
Math
70
Physics
70
Biology
70
Chemistry
70
Reasoning
60
General
60
Factuality
40
Frontend Development
40
Code
40

Pricing

Input Price$16.5 / 1M tokens
Output Price$66 / 1M tokens
Blended Price (3:1)$28.875 / 1M tokens
Cache Read Price$7.5 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

1 providers

ProviderInputOutput
1OpenAIPRIMARY
$16.5
$66

Compare pricing across different API providers for this model.

External Sources