メインコンテンツへスキップ

o3-mini

OpenAIOpenAI o-seriesProprietary

説明

A smaller variant of O3, expected to offer enhanced multimodal capabilities, improved reasoning, and more efficient resource utilization compared to previous models while maintaining strong performance on core tasks.

リリース日
2025-01-31
パラメータ
コンテキスト長
200K
モダリティ
text

能力レーダー

35
general
65
coding
83
reasoning
49
science
40
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング260
44.0
AA
総合ランキング286
44.0
AA
科学218
52.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)77.2%自己申告

Code

Aider-Polyglot66.7%自己申告
Aider-Polyglot Edit60.4%自己申告
SWE-Bench Verified49.3%自己申告
SWE-Lancer18.0%自己申告
SWE-Lancer (IC-Diamond subset)7.4%自己申告

Communication

Multi-IF79.5%自己申告
TAU-bench Retail57.6%自己申告
Multi-Challenge39.9%自己申告
TAU-bench Airline32.4%自己申告

Factuality

SimpleQA15.0%自己申告

Finance

MMLU86.9%自己申告

General

IFEvalGoogle Research (2023)93.9%自己申告
LiveBench84.6%自己申告
Multilingual MMLU80.7%自己申告
Internal API instruction following (hard)50.0%自己申告

Language

COLLIE98.7%自己申告

Long Context

OpenAI-MRCR: 2 needle 128k18.7%自己申告
ComplexFuncBench17.6%自己申告

Math

MATH97.9%自己申告
MGSM92.0%自己申告
AIME 202487.3%自己申告
FrontierMath9.2%自己申告

Reasoning

Graphwalks parents <128k58.3%自己申告
Graphwalks BFS <128k51.0%自己申告

AA評価指数

(Artificial Analysis)
Intelligence Index(Artificial Analysis)
19.2
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
1.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Aime(MAA (Mathematical Association of America))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.4
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.3
Hle(Center for AI Safety + Scale AI (2025))
0.1
Terminalbench Hard(Stanford × Laude Institute (2026))
0.1

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Writing
100
Legal
90
Instruction Following
90
Language
90
Finance
90
Healthcare
90
Math
80
Physics
80
Biology
80
Chemistry
80
Reasoning
60
Structured Output
60
General
60
Spatial Reasoning
50
Frontend Development
50
Communication
50
Code
40
Tool Calling
40
Long Context
20
Factuality
10

価格設定

入力価格$1.1 / 1Mトークン
出力価格$4.4 / 1Mトークン
混合価格(3:1)$1.925 / 1Mトークン
キャッシュ読み取り価格$0.55 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

7 プロバイダー

最安: OpenAI最高: Azure
プロバイダー入力出力
1OpenAIプライマリ
$1.1
$4.4
2Abacus
$1.1
$4.4
3Jiekou.AI
$1.1
$4.4
4Helicone
$1.1
$4.4
5Azure Cognitive Services
$1.1
$4.4
6LLM Gateway
$1.1
$4.4
7Azure
$1.1
$4.4

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク