Inkling Small
Описание
Inkling-Small is Thinking Machines Lab's efficient open-weights MoE multimodal model (276B total / 12B active parameters) released under Apache 2.0. It accepts text, image, and audio inputs and generates text, with native reasoning, variable thinking effort, and a context window up to 1M tokens (Tinker exposes 64K and 256K configurations). Hugging Face weights: thinkingmachines/Inkling-Small and thinkingmachines/Inkling-Small-NVFP4. Vendor self-host VRAM: BF16 ≥ ~600 GB aggregated; NVFP4 ≥ ~180 GB aggregated. SWE-Bench Verified 80.2% vs Inkling 77.6% (same bash-only harness). Fine-tuning and playground chat are available via Tinker. Official Tinker serverless inference (256K, list): $0.30 / $1.20 per 1M input/output tokens ($0.06 cached).
Радар способностей
Рейтинги
| Домен | #Место | Оценка | Источник |
|---|---|---|---|
| Агентные возможности | 28 | 53.0 | LS |
| Рейтинг кодинга | 121 | 72.0 | AA |
| Общий рейтинг | 191 | 56.0 | AA |
| Мультимодальный рейтинг | 67 | 49.0 | LS |
| Наука | 89 | 73.0 | AA |
Оценки бенчмарков (LLM Stats)
(LLM Stats (zeroeval))Agents
Audio
Chat
Factuality
General
Instruction Following
Math
Reasoning
Vision
Индексы оценки AA
(Artificial Analysis)Оценки категорий LLM Stats
(LLM Stats (zeroeval))Цены
Скорость
Рейтинг цен провайдеров
Рейтинг цен провайдеров
8 провайдеров
Сравнение цен разных API-провайдеров для этой модели.