メインコンテンツへスキップ

MiniMax M3

MiniMaxMiniMaxオープンウエイトMIT · 商用利用可

説明

MiniMax M3 is the first open-weight model to combine three frontier capabilities: top-tier coding and agentic performance, a 1M-token context window, and native multimodality. It is powered by MiniMax Sparse Attention (MSA), a new sparse attention architecture that partitions the KV cache into blocks to cut per-token compute at long context — roughly 1/20 the cost of the previous generation at 1M tokens, with more than 9x faster prefill and more than 15x faster decode while matching full attention on most capabilities. Trained with mixed-modality data from step zero across 100T+ tokens, M3 natively supports image and video input and can operate a desktop computer. On SWE-Bench Pro it scores 59.0%, surpassing GPT-5.5 and Gemini 3.1 Pro and approaching Opus 4.7, and on BrowseComp it scores 83.5%, surpassing Opus 4.7. M3 supports toggling thinking on or off at request time.

リリース日
2026-06-01
パラメータ
コンテキスト長
1.0M
モダリティ
image, text, video

能力レーダー

70
general
60
coding
18
reasoning
68
science推定
80
agents
80
multimodal

専門的な科学ベンチマークが利用できない場合、Scienceは推論プロキシを使用して推定します。

ランキング

ドメイン#順位スコアソース
エージェント能力65
46.0
LS
数学的推論31
77.0
LB
マルチモーダルランキング11
64.0
LS
推論27
74.0
LB

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

YC-Bench2100000.00 / 10000000自己申告
GDPval-AA1431.00 / 3000
SpreadSheetBench-v189.3%自己申告
BrowseCompOpenAI (2025)83.5%自己申告
BankerToolBench76.1%自己申告
GDPval-Rubrics74.8%自己申告
Claw-Eval74.5%自己申告
MCP Atlas74.2%自己申告
DRACO73.2%自己申告
OSWorld-Verified70.1%自己申告
Terminal-Bench 2.166.0%自己申告
SVG-Bench63.7%自己申告
SWE-Bench ProPrinceton NLP (2024)59.0%自己申告
PaperBench52.6%自己申告
VIBE-V250.1%自己申告
LOCA-Bench (256k)49.3%自己申告
Finance Agent v248.3%
OfficeQA Pro45.1%自己申告
NL2Repo42.1%自己申告
LiveSQLBench40.2%自己申告
SWE Atlas - Codebase QnA37.9%自己申告
PostTrainBench37.1%自己申告
SWE-fficiency34.8%自己申告
SWE Atlas - Test Writing30.8%自己申告
KernelBench Hard28.8%自己申告
APEX-Agents27.7%自己申告
CL-bench20.5%自己申告
FrontierCode 1.114.7%

Code

SWE-Bench Verified80.5%自己申告

General

MMMU-Pro78.1%自己申告
LiveBench70.0%

Healthcare

VideoMMMU84.6%自己申告

Math

USAMO 202636.00 / 42自己申告
IMO 202535.00 / 42自己申告

Multimodal

OmniDocBench 1.591.6%自己申告
Video-MME85.4%自己申告

AA評価指数

(Artificial Analysis)

AA評価データがありません

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Legal
100
Finance
100
Agents
100
Reasoning
87
General
70
Math
18
Productivity
90
Structured Output
90
Multimodal
80
Search
80
Frontend Development
80
Healthcare
80
Tool Calling
80
Vision
80
Code
60
Systems
40

価格設定

入力価格$0 / 1Mトークン
出力価格$0 / 1Mトークン
混合価格(3:1)$0 / 1Mトークン
キャッシュ読み取り価格$0.06 / 1Mトークン

速度

速度データがありません

プロバイダー価格ランキング

プロバイダー価格ランキング

7 プロバイダー

最安: MiniMax最高: Wafer
プロバイダー入力出力
1MiniMaxプライマリ
$0
$0
2Novita
$0
$0
3Together
$0
$0
4Fireworks
$0
$0
5MiniMax (minimax.io)
$0.3
$1.2
6MiniMax (minimaxi.com)
$0.3
$1.2
7Wafer
$0.33
$1.32

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク