メインコンテンツへスキップ

o3

OpenAIOpenAI o-seriesProprietary

説明

OpenAI's most powerful reasoning model. o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following. Use it to think through multi-step problems that involve analysis across text, code, and images.

リリース日
2025-04-16
パラメータ
コンテキスト長
200K
モダリティ
image, pdf, text

能力レーダー

46
general
72
coding
90
reasoning
56
science
70
agents
85
multimodal

ランキング

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

Tau-bench63.0%自己申告
BrowseCompOpenAI (2025)49.7%自己申告

Biology

GPQANYU + Cohere + Anthropic (2023)83.3%自己申告

Code

Aider-Polyglot81.3%自己申告
SWE-Bench Verified69.1%自己申告

Communication

Tau2 Retail80.2%自己申告
Tau2 Airline64.8%自己申告
Multi-Challenge60.4%自己申告
Tau2 Telecom58.2%自己申告

General

MMMU82.9%自己申告
MMMU-Pro76.4%自己申告

Healthcare

VideoMMMU83.3%自己申告

Language

COLLIE98.4%自己申告

Math

AIME 202491.6%自己申告
MathVista86.8%自己申告
AIME 202586.4%自己申告
FrontierMath15.8%自己申告
Humanity's Last Exam14.7%自己申告

Multimodal

CharXiv-R78.6%自己申告

Reasoning

ARC-AGI88.0%自己申告
ERQA64.0%自己申告
ARC-AGI v26.5%

AA評価指数

(Artificial Analysis)
Math Index(Artificial Analysis)
88.3
Intelligence Index(Artificial Analysis)
31.1
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
1.0
Aime(MAA (Mathematical Association of America))
0.9
Aime 25(MAA (Mathematical Association of America))
0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.9
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.8
Lcr(Artificial Analysis)
0.7
Ifbench(Google Research (2023))
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.4
Terminalbench Hard(Stanford × Laude Institute (2026))
0.4
Hle(Center for AI Safety + Scale AI (2025))
0.2

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Language
100
Writing
100
Multimodal
80
Physics
80
Healthcare
80
Biology
80
Chemistry
80
Code
80
Reasoning
70
Frontend Development
70
General
70
Communication
70
Tool Calling
70
Math
60
Agents
60
Vision
60
Search
50
Spatial Reasoning
50

価格設定

入力価格$2 / 1Mトークン
出力価格$8 / 1Mトークン
混合価格(3:1)$3.5 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン

速度

トークン/秒122.2
初トークン遅延7.63s
初回答遅延7.63s

プロバイダー価格ランキング

プロバイダー価格ランキング

16 プロバイダー

最安: Poe最高: Jiekou.AI
プロバイダー入力出力
1Poe最安
$1.8
$7.2
2OpenAIプライマリ
$2
$8
3NanoGPT
$2
$8
4Abacus
$2
$8
5OpenRouter
$2
$8
6Kilo Gateway
$2
$8
7Cloudflare AI Gateway
$2
$8
8Helicone
$2
$8
9Azure Cognitive Services
$2
$8
10Vercel AI Gateway
$2
$8
11LLM Gateway
$2
$8
12Azure
$2
$8
13NEAR AI Cloud
$2
$8
14Merge Gateway
$2
$8
15Impossibl
$2
$8
16Jiekou.AI
$10
$40

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク