Skip to main content

MiniMax-M3

MiniMaxMiniMaxOpen WeightMIT · Commercial OK

Description

MiniMax M3 is the first open-weight model to combine three frontier capabilities: top-tier coding and agentic performance, a 1M-token context window, and native multimodality. It is powered by MiniMax Sparse Attention (MSA), a new sparse attention architecture that partitions the KV cache into blocks to cut per-token compute at long context — roughly 1/20 the cost of the previous generation at 1M tokens, with more than 9x faster prefill and more than 15x faster decode while matching full attention on most capabilities. Trained with mixed-modality data from step zero across 100T+ tokens, M3 natively supports image and video input and can operate a desktop computer. On SWE-Bench Pro it scores 59.0%, surpassing GPT-5.5 and Gemini 3.1 Pro and approaching Opus 4.7, and on BrowseComp it scores 83.5%, surpassing Opus 4.7. M3 supports toggling thinking on or off at request time.

Release Date
2026-06-01
Parameters
428.0B
Context Length
1.0M
Modalities
image, text, video

Capability Radar

44
general
57
coding
93
reasoning
65
science
80
agents
80
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability50
47.0
LS
Code Ranking77
78.0
AA
General Ranking39
83.0
AA
Math Reasoning41
77.0
LB
Multimodal Ranking26
61.0
LS
Reasoning37
74.0
LB
Science57
81.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

YC-Bench2100000.00 / 10000000SR
SpreadSheetBench-v189.3%SR
BankerToolBench76.1%SR
DRACO73.2%SR
LOCA-Bench (256k)49.3%SR
Finance Agent v248.3%
OfficeQA Pro45.1%SR
PostTrainBench37.1%SR

Code

Claw-Eval74.5%SR
SVG-Bench63.7%SR
PaperBench52.6%SR
VIBE-V250.1%SR
NL2Repo42.1%SR
LiveSQLBench40.2%SR
SWE Atlas - Codebase QnA37.9%SR
SWE-fficiency34.8%SR
SWE Atlas - Test Writing30.8%SR
KernelBench Hard28.8%SR
CL-bench20.5%SR

General

GDPval-Rubrics74.8%SR

Math

USAMO 202636.00 / 42SR
IMO 202535.00 / 42SR
LiveBench70.0%

Multimodal

Video-MME85.4%SR
VideoMMMU84.6%SR
OSWorld-Verified70.1%SR

Reasoning

BrowseCompOpenAI (2025)83.5%SR
SWE-Bench Verified80.5%SR
MCP Atlas74.2%SR
Terminal-Bench 2.166.0%SR
SWE-Bench ProPrinceton NLP (2024)59.0%SR
APEX-Agents27.7%SR
FrontierCode 1.114.7%

Vision

OmniDocBench 1.591.6%SR
MMMU-Pro78.1%SR

AA Evaluation Indices

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
92.9
Tau2(Sierra + U Toronto + Vector Institute (2025))
88.9
Ifbench(Google Research (2023))
82.9
Lcr(Artificial Analysis)
80.3
Terminalbench V2 1
65.2
Coding Index(Artificial Analysis)
58.6
Intelligence Index(Artificial Analysis)
45.4
Scicode(UIUC + Argonne National Lab (2024))
45.4
Terminalbench Hard(Stanford × Laude Institute (2026))
42.4
Hle(Center for AI Safety + Scale AI (2025))
39.0
Tau Banking
15.3

LLM Stats Category Scores

(LLM Stats (zeroeval))
Math
18
Reasoning
3
General
2
Productivity
90
Structured Output
90
Multimodal
80
Search
80
Frontend Development
80
Healthcare
80
Tool Calling
80
Vision
80
Agents
60
Code
60
Finance
50
Systems
40

Pricing

Input Price$0.3 / 1M tokens
Output Price$1.2 / 1M tokens
Blended Price (3:1)$0.525 / 1M tokens
Cache Read Price$0.06 / 1M tokens

Speed

Tokens/sec114.4
Time to First Token0.95s
Time to Answer18.42s

Provider Price Ranking

Provider Price Ranking

27 providers

Cheapest: MiniMaxMost Expensive: LLM Gateway
ProviderInputOutput
1MiniMaxCheapest
$0
$0
2Fireworks
$0
$0
3Novita
$0
$0
4Together
$0
$0
5EmpirioLabs AI
$0.225
$0.9
6Vancine
$0.24
$0.96
7NanoGPT
$0.3
$1.2
8OpenRouter
$0.3
$1.2
9OpenCode Go
$0.3
$1.2
10Kilo Gateway
$0.3
$1.2
11OpenCode Zen
$0.3
$1.2
12Requesty
$0.3
$1.2
13Vercel AI Gateway
$0.3
$1.2
14MiniMax (minimax.io)
$0.3
$1.2
15DevPass (LLM Gateway)
$0.3
$1.2
16MiniMax (minimaxi.com)
$0.3
$1.2
17OrcaRouter
$0.3
$1.2
18Merge Gateway
$0.3
$1.2
19Jalapeno Cloud
$0.3
$1.2
20Charm Hyper
$0.32664
$1.30656
21Wafer
$0.33
$1.32
22CrossModel
$0.33
$1.32
23Cortecs
$0.395
$1.977
24TensorX
$0.4
$2
25ZenMux
$0.6
$2.4
26Ofox
$0.6
$2.4
27LLM Gateway
$0.6
$2.4

Compare pricing across different API providers for this model.

External Sources