Saltar al contenido principal

Kimi K2.6 (Non-reasoning)

KimiKimiOpen WeightModified MIT License

Descripción

Kimi K2.6 is Moonshot AI's open-source, native multimodal agentic model focused on state-of-the-art coding, long-horizon execution, and agent swarm capabilities. It scales horizontally to 300 sub-agents executing 4,000 coordinated steps, dynamically decomposing tasks into parallel, domain-specialized subtasks. K2.6 unifies text, image, and video input with thinking and non-thinking modes, supports a 256K context, and powers proactive 24/7 background agents that manage schedules, execute code, and orchestrate cross-platform operations without human oversight.

Fecha de lanzamiento
2026-04-20
Parámetros
1.0T
Longitud del contexto
262K
Modalidades
image, text, video

Radar de capacidades

23
general
60
coding
79
reasoning
61
science
60
agents
80
multimodal

Rankings

Dominio#PosiciónPuntuaciónFuente
Capacidad agéntica35
52.0
LS
Ranking de codificación196
68.0
AA
Ranking general198
54.0
AA
Razonamiento matemático46
84.0
LB
Ranking multimodal51
57.0
LS
Razonamiento45
79.0
LB
Ciencia223
56.0
AA

Puntuaciones de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

Agents

MCP-Mark55.9%Aut.
Toolathlon50.0%Aut.
Finance Agent v244.9%

Code

Claw-Eval80.9%Aut.
FrontierSWE27.0%

Math

AIME 202696.4%Aut.
MathVision93.2%Aut.
HMMT Feb 2692.7%Aut.
IMO-AnswerBench86.0%Aut.
LiveBench72.2%

Multimodal

OSWorld-Verified73.1%Aut.

Reasoning

GPQANYU + Cohere + Anthropic (2023)90.5%Aut.
LiveCodeBench v689.6%Aut.
CharXiv-R86.7%Aut.
BrowseCompOpenAI (2025)86.3%Aut.
SWE-Bench Verified80.2%Aut.
SWE-bench Multilingual76.7%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)66.7%Aut.
OJBench60.6%Aut.
SWE-Bench ProPrinceton NLP (2024)58.6%Aut.
SciCode52.2%Aut.
Humanity's Last Exam36.4%Aut.
APEX-Agents27.9%Aut.

Search

DeepSearchQA83.0%Aut.
WideSearch80.8%Aut.

Vision

V*96.9%Aut.
MMMU-Pro80.1%Aut.
BabyVision68.5%Aut.

Índices de evaluación AA

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
93.9
Gpqa(NYU + Cohere + Anthropic (2023))
78.8
Lcr(Artificial Analysis)
69.7
Ifbench(Google Research (2023))
44.3
Terminalbench Hard(Stanford × Laude Institute (2026))
37.9
Intelligence Index(Artificial Analysis)
23.6
Hle(Center for AI Safety + Scale AI (2025))
19.6

Puntuaciones por categoría LLM Stats

(LLM Stats (zeroeval))
Math
80
Multimodal
80
Search
80
Frontend Development
80
Vision
80
Physics
70
Reasoning
70
General
70
Biology
70
Chemistry
70
Agents
60
Code
60
Tool Calling
60
Finance
40

Precios

Precio de entrada$0.95 / 1M tokens
Precio de salida$4 / 1M tokens
Precio mixto (3:1)$1.712 / 1M tokens
Precio de lectura caché$0.15 / 1M tokens

Velocidad

Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s

Ranking de Precios por Proveedor

Ranking de Precios por Proveedor

17 proveedores

Más barato: DeepInfraMás caro: TensorX
ProveedorEntradaSalida
1DeepInfraMás barato
$0
$0
2Novita
$0
$0
3Fireworks
$0
$0
4Moonshot AI
$0
$0
5NanoGPT
$0.5
$2.6
6OpenRouter
$0.65
$3.41
7Kilo Gateway
$0.65
$3.41
8Lilac
$0.7
$3.5
9FastRouter
$0.75
$3.5
10NovitaAI
$0.8
$3.4
11KimiPRINCIPAL
$0.95
$4
12ZenMux
$0.95
$4
13Vercel AI Gateway
$0.95
$4
14Ambient
$0.95
$4
15Ofox
$0.95
$4
16TokenGo
$0.95
$4
17TensorX
$1
$4

Comparar precios entre diferentes proveedores de API para este modelo.

Fuentes externas