o1-preview
OpenAIOpenAI o-seriesProprietary
説明
A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.
リリース日
2024-09-12
パラメータ
—
コンテキスト長
200K
モダリティ
image, pdf, text
能力レーダー
37
general
34
coding
82
reasoning
77
science
68
agents
80
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 357 | 42.0 | AA |
| 総合ランキング | 331 | 41.0 | AA |
| 科学 | 84 | 77.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Factuality
SimpleQA
42.4%自己申告
General
MMLU
90.8%自己申告
Math
MGSM
90.8%自己申告
MATH
85.5%自己申告
LiveBench
52.3%自己申告
AIME 2024
42.0%自己申告
Reasoning
GPQANYU + Cohere + Anthropic (2023)
73.3%自己申告
SWE-Bench Verified
41.3%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))92.4
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))84.8
Math Index(Artificial Analysis)79.7
Aime 25(MAA (Mathematical Association of America))79.7
Gpqa(NYU + Cohere + Anthropic (2023))76.5
Coding Index(Artificial Analysis)34.0
Intelligence Index(Artificial Analysis)11.4
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Language90
Legal90
Finance90
Healthcare90
Math70
Physics70
Biology70
Chemistry70
Reasoning60
General60
Factuality40
Frontend Development40
Code40
価格設定
入力価格$16.5 / 1Mトークン
出力価格$66 / 1Mトークン
混合価格(3:1)$28.875 / 1Mトークン
キャッシュ読み取り価格$7.5 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
1 プロバイダー
プロバイダー入力出力
1OpenAIプライマリ
$16.5
$66
このモデルの異なるAPIプロバイダー間の価格を比較。