メインコンテンツへスキップ

o3

OpenAIOpenAI o-seriesProprietary

説明

OpenAI's most powerful reasoning model. o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following. Use it to think through multi-step problems that involve analysis across text, code, and images.

リリース日
2025-04-16
パラメータ
—
コンテキスト長
200K
モダリティ
image, pdf, text

能力レーダー

40
general
81
coding
90
reasoning
63
science
70
agents
85
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング133
76.0
AA
総合ランキング112
65.0
AA
マルチモーダルランキング42
59.0
LS
科学204
58.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Chat

Tau2 Airline64.8%自己申告
Multi-Challenge60.4%自己申告

Communication

Tau2 Retail80.2%自己申告
Tau2 Telecom58.2%自己申告

General

Aider-Polyglot81.3%自己申告
Tau-bench63.0%自己申告

Language

COLLIE98.4%自己申告

Math

AIME 202491.6%自己申告
MathVista86.8%自己申告
AIME 202586.4%自己申告
FrontierMath15.8%自己申告

Multimodal

VideoMMMU83.3%自己申告
MMMU82.9%自己申告

Reasoning

ARC-AGI88.0%自己申告
GPQANYU + Cohere + Anthropic (2023)83.3%自己申告
CharXiv-R78.6%自己申告
SWE-Bench Verified69.1%自己申告
BrowseCompOpenAI (2025)49.7%自己申告
Humanity's Last Exam14.7%自己申告
ARC-AGI v26.5%

Vision

MMMU-Pro76.4%自己申告
ERQA64.0%自己申告

AA評価指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
99.2
Aime(MAA (Mathematical Association of America))
90.3
Aime 25(MAA (Mathematical Association of America))
88.3
Math Index(Artificial Analysis)
88.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
85.3
Gpqa(NYU + Cohere + Anthropic (2023))
82.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
80.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
80.7
Lcr(Artificial Analysis)
74.7
Ifbench(Google Research (2023))
71.4
Terminalbench Hard(Stanford × Laude Institute (2026))
37.1
Intelligence Index(Artificial Analysis)
20.2
Hle(Center for AI Safety + Scale AI (2025))
20.1

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Language
100
Writing
100
Multimodal
80
Physics
80
Healthcare
80
Biology
80
Chemistry
80
Code
80
Reasoning
70
Frontend Development
70
General
70
Communication
70
Tool Calling
70
Chat
60
Math
60
Agents
60
Vision
60
Search
50
Spatial Reasoning
50

価格設定

入力価格$2 / 1Mトークン
出力価格$8 / 1Mトークン
混合価格(3:1)$3.5 / 1Mトークン
キャッシュ読み取り価格$0.5 / 1Mトークン

速度

トークン/秒166.3
初トークン遅延5.31s
初回答遅延5.31s

プロバイダー価格ランキング

プロバイダー価格ランキング

20 プロバイダー

最安: Poe最高: Jiekou.AI
プロバイダー入力出力
1Poe最安
$1.8
$7.2
2OpenAIプライマリ
$2
$8
3302.AI
$2
$8
4NanoGPT
$2
$8
5Abacus
$2
$8
6OpenRouter
$2
$8
7Kilo Gateway
$2
$8
8Cloudflare AI Gateway
$2
$8
9Helicone
$2
$8
10AIHubMix
$2
$8
11Azure Cognitive Services
$2
$8
12Vercel AI Gateway
$2
$8
13DevPass (LLM Gateway)
$2
$8
14Azure
$2
$8
15NEAR AI Cloud
$2
$8
16Merge Gateway
$2
$8
17Impossibl
$2
$8
18Eden AI
$2
$8
19LLM Gateway
$2
$8
20Jiekou.AI
$10
$40

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク