o3-mini
OpenAIOpenAI o-seriesProprietary
説明
A smaller variant of O3, expected to offer enhanced multimodal capabilities, improved reasoning, and more efficient resource utilization compared to previous models while maintaining strong performance on core tasks.
リリース日
2025-01-31
パラメータ
—
コンテキスト長
200K
モダリティ
text
能力レーダー
35
general
65
coding
83
reasoning
49
science
40
agents
0
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 260 | 44.0 | AA |
| 総合ランキング | 286 | 44.0 | AA |
| 科学 | 218 | 52.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
77.2%自己申告
Code
Aider-Polyglot
66.7%自己申告
Aider-Polyglot Edit
60.4%自己申告
SWE-Bench Verified
49.3%自己申告
SWE-Lancer
18.0%自己申告
SWE-Lancer (IC-Diamond subset)
7.4%自己申告
Communication
Multi-IF
79.5%自己申告
TAU-bench Retail
57.6%自己申告
Multi-Challenge
39.9%自己申告
TAU-bench Airline
32.4%自己申告
Factuality
SimpleQA
15.0%自己申告
Finance
MMLU
86.9%自己申告
General
IFEvalGoogle Research (2023)
93.9%自己申告
LiveBench
84.6%自己申告
Multilingual MMLU
80.7%自己申告
Internal API instruction following (hard)
50.0%自己申告
Language
COLLIE
98.7%自己申告
Long Context
OpenAI-MRCR: 2 needle 128k
18.7%自己申告
ComplexFuncBench
17.6%自己申告
Math
MATH
97.9%自己申告
MGSM
92.0%自己申告
AIME 2024
87.3%自己申告
FrontierMath
9.2%自己申告
Reasoning
Graphwalks parents <128k
58.3%自己申告
Graphwalks BFS <128k
51.0%自己申告
AA評価指数
(Artificial Analysis)Intelligence Index(Artificial Analysis)19.2
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))1.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Aime(MAA (Mathematical Association of America))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.7
Scicode(UIUC + Argonne National Lab (2024))0.4
Tau2(Sierra + U Toronto + Vector Institute (2025))0.3
Hle(Center for AI Safety + Scale AI (2025))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Writing100
Legal90
Instruction Following90
Language90
Finance90
Healthcare90
Math80
Physics80
Biology80
Chemistry80
Reasoning60
Structured Output60
General60
Spatial Reasoning50
Frontend Development50
Communication50
Code40
Tool Calling40
Long Context20
Factuality10
価格設定
入力価格$1.1 / 1Mトークン
出力価格$4.4 / 1Mトークン
混合価格(3:1)$1.925 / 1Mトークン
キャッシュ読み取り価格$0.55 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
7 プロバイダー
最安: OpenAI最高: Azure
プロバイダー入力出力
1OpenAIプライマリ
$1.1
$4.4
2Abacus
$1.1
$4.4
3Jiekou.AI
$1.1
$4.4
4Helicone
$1.1
$4.4
5Azure Cognitive Services
$1.1
$4.4
6LLM Gateway
$1.1
$4.4
7Azure
$1.1
$4.4
このモデルの異なるAPIプロバイダー間の価格を比較。