Kimi K2.6 (Non-reasoning)
KimiKimiOpen WeightModified MIT License
Description
Kimi K2.6 is Moonshot AI's open-source, native multimodal agentic model focused on state-of-the-art coding, long-horizon execution, and agent swarm capabilities. It scales horizontally to 300 sub-agents executing 4,000 coordinated steps, dynamically decomposing tasks into parallel, domain-specialized subtasks. K2.6 unifies text, image, and video input with thinking and non-thinking modes, supports a 256K context, and powers proactive 24/7 background agents that manage schedules, execute code, and orchestrate cross-platform operations without human oversight.
Release Date
2026-04-20
Parameters
1.0T
Context Length
262K
Modalities
image, text, video
Capability Radar
32
general
40
coding
79
reasoning
53
science
60
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 28 | 56.0 | LS |
| Code Ranking | 130 | 68.0 | AA |
| General Ranking | 142 | 63.0 | AA |
| Math Reasoning | 27 | 84.0 | LB |
| Multimodal Ranking | 14 | 63.0 | LS |
| Reasoning | 26 | 79.0 | LB |
| Science | 168 | 60.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
BrowseCompOpenAI (2025)
86.3%SR
DeepSearchQA
83.0%SR
Claw-Eval
80.9%SR
WideSearch
80.8%SR
OSWorld-Verified
73.1%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)
66.7%SR
SWE-Bench ProPrinceton NLP (2024)
58.6%SR
MCP-Mark
55.9%SR
Toolathlon
50.0%SR
Finance Agent v2
44.9%
APEX-Agents
27.9%SR
FrontierSWE
27.0%
Biology
GPQANYU + Cohere + Anthropic (2023)
90.5%SR
SciCode
52.2%SR
Code
SWE-Bench Verified
80.2%SR
SWE-bench Multilingual
76.7%SR
General
LiveCodeBench v6
89.6%SR
MMMU-Pro
80.1%SR
LiveBench
72.2%
Math
AIME 2026
96.4%SR
MathVision
93.2%SR
HMMT Feb 26
92.7%SR
IMO-AnswerBench
86.0%SR
Humanity's Last Exam
36.4%SR
Multimodal
V*
96.9%SR
CharXiv-R
86.7%SR
BabyVision
68.5%SR
Reasoning
OJBench
60.6%SR
AA Evaluation Indices
(Artificial Analysis)Intelligence Index(Artificial Analysis)35.4
Tau2(Sierra + U Toronto + Vector Institute (2025))0.9
Gpqa(NYU + Cohere + Anthropic (2023))0.8
Lcr(Artificial Analysis)0.7
Ifbench(Google Research (2023))0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Terminalbench Hard(Stanford × Laude Institute (2026))0.4
Hle(Center for AI Safety + Scale AI (2025))0.2
LLM Stats Category Scores
(LLM Stats (zeroeval))Math80
Multimodal80
Search80
Frontend Development80
Vision80
Physics70
Reasoning70
General70
Biology70
Chemistry70
Agents60
Code60
Tool Calling60
Finance40
Pricing
Input Price$0.95 / 1M tokens
Output Price$4 / 1M tokens
Blended Price (3:1)$1.712 / 1M tokens
Cache Read Price$0.0944 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
17 providers
Cheapest: DeepInfraMost Expensive: TensorX
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2Fireworks
$0
$0
3Novita
$0
$0
4Moonshot AI
$0
$0
5Together
$0
$0
6NanoGPT
$0.5
$2.6
7OpenRouter
$0.5605
$2.36
8Lilac
$0.7
$3.5
9FastRouter
$0.75
$3.5
10NovitaAI
$0.8
$3.4
11Kilo Gateway
$0.8
$3.4
12KimiPRIMARY
$0.95
$4
13ZenMux
$0.95
$4
14Vercel AI Gateway
$0.95
$4
15Ambient
$0.95
$4
16Ofox
$0.95
$4
17TensorX
$1
$4
Compare pricing across different API providers for this model.