Kimi K3 (max)
KimiKimiOpen WeightKimi K3 License · Commercial OK
Description
Kimi K3 is Moonshot AI's flagship open Mixture-of-Experts model for long-horizon coding, knowledge work, and reasoning. It has 2.8 trillion total parameters, activates 16 of 896 experts, and combines Kimi Delta Attention, Attention Residuals, and Stable LatentMoE. The model supports a 1M-token context window, native image and video understanding, tool calling, structured output, and always-on reasoning with max thinking effort at launch.
Release Date
2026-07-16
Parameters
2.8T
Context Length
1.0M
Modalities
image, text, video
Capability Radar
57
general
74
coding
93
reasoning
72
science
70
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 9 | 63.0 | LS |
| Code Ranking | 2 | 98.0 | AA |
| General Ranking | 6 | 95.0 | AA |
| Multimodal Ranking | 7 | 70.0 | LS |
| Science | 8 | 93.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
AA-Briefcase
1548.00 / 3000
Program Bench
77.8%SR
DECK-Bench
73.5%SR
Toolathlon
73.2%SR
Kimi Code Bench v2
72.9%SR
OfficeQA Pro
63.3%SR
Job Bench
52.9%SR
MLS-Bench Lite
48.3%SR
PostTrainBench
36.6%SR
SpreadsheetBench 2
34.8%SR
AutomationBench
30.8%SR
Code
FrontierSWE
81.2%SR
DeepSWE 1.1
69.0%
DeepSWE
67.5%SR
SWE-Marathon
42.0%SR
Math
MathVision
97.8%SR
Reasoning
GPQANYU + Cohere + Anthropic (2023)
93.5%SR
CharXiv-R
91.3%SR
BrowseCompOpenAI (2025)
91.2%SR
Terminal-Bench 2.1
88.3%SR
MCP Atlas
84.2%SR
Humanity's Last Exam
56.0%SR
APEX-Agents
37.6%SR
Search
DeepSearchQA
95.0%SR
Vision
OmniDocBench
91.1%SR
BabyVision
85.7%SR
MMMU-Pro (with tools)
83.4%SR
MMMU-Pro
81.6%SR
PerceptionBench
58.5%SR
WorldVQA
51.0%SR
ZEROBench
0.41 / 100SR
AA Evaluation Indices
(Artificial Analysis)Gpqa(NYU + Cohere + Anthropic (2023))93.5
Terminalbench V2 185.0
Lcr(Artificial Analysis)82.7
Coding Index(Artificial Analysis)76.2
Intelligence Index(Artificial Analysis)59.7
Scicode(UIUC + Argonne National Lab (2024))58.7
Hle(Center for AI Safety + Scale AI (2025))46.9
Tau Banking46.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Productivity100
Agents92
Reasoning78
General60
Physics90
Search90
Biology90
Chemistry90
Math80
Multimodal80
Code70
Tool Calling70
Vision70
Systems40
Pricing
Input Price$3 / 1M tokens
Output Price$15 / 1M tokens
Blended Price (3:1)$6 / 1M tokens
Cache Read Price$0.3 / 1M tokens
Speed
Tokens/sec36.4
Time to First Token3.56s
Time to Answer58.53s
Provider Price Ranking
Provider Price Ranking
25 providers
Cheapest: FireworksMost Expensive: Tinfoil
ProviderInputOutput
1FireworksCheapest
$0
$0.00002
2Moonshot AI
$0
$0.00002
3Together
$0
$0.00002
4Novita
$0
$0.00002
5CrofAI
$2
$8
6Requesty
$2.25
$11.25
7Vancine
$2.4
$12
8DevPass (LLM Gateway)
$2.83
$14.13
9DigitalOcean
$2.85
$14.25
10KimiPRIMARY
$3
$15
11OpenCode Go
$3
$15
12Vivgrid
$3
$15
13GitHub Copilot
$3
$15
14OpenCode Zen
$3
$15
15AIHubMix
$3
$15
16Moonshot AI (China)
$3
$15
17Cortecs
$3
$14.999
18Neuralwatt
$3
$15
19Neon
$3
$15
20EmpirioLabs AI
$3
$15
21CoralBricks
$3
$15
22Charm Hyper
$3.2664
$16.332
23Venice AI
$3.75
$18.75
24GreenPT
$3.762
$18.81
25Tinfoil
$4
$20
Compare pricing across different API providers for this model.