메인 콘텐츠로 건너뛰기

Inkling Small

Thinking Machines오픈 웨이트Apache 2.0 · 상업적 사용 가능

설명

Inkling-Small is Thinking Machines Lab's efficient open-weights MoE multimodal model (276B total / 12B active parameters) released under Apache 2.0. It accepts text, image, and audio inputs and generates text, with native reasoning, variable thinking effort, and a context window up to 1M tokens (Tinker exposes 64K and 256K configurations). Hugging Face weights: thinkingmachines/Inkling-Small and thinkingmachines/Inkling-Small-NVFP4. Vendor self-host VRAM: BF16 ≥ ~600 GB aggregated; NVFP4 ≥ ~180 GB aggregated. SWE-Bench Verified 80.2% vs Inkling 77.6% (same bash-only harness). Fine-tuning and playground chat are available via Tinker. Official Tinker serverless inference (256K, list): $0.30 / $1.20 per 1M input/output tokens ($0.06 cached).

출시일
2026-07-30
파라미터
276.0B
컨텍스트 길이
1.0M
모달리티
audio, image, text

능력 레이더

33
general
52
coding
90
reasoning
64
science
50
agents
80
multimodal

랭킹

도메인#순위점수소스
에이전트형 역량28
53.0
LS
코딩 랭킹121
72.0
AA
종합 랭킹191
56.0
AA
멀티모달 랭킹67
49.0
LS
과학89
73.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Agents

AA-Briefcase917.00 / 3000자체 보고
Toolathlon54.4%자체 보고
Tau3 Banking15.5%자체 보고

Audio

MMAU77.0%자체 보고

Chat

VoiceBench Avg90.1%자체 보고

Factuality

SimpleQA Verified20.6%자체 보고

General

GDPval-AA1269.00 / 3000자체 보고
Global-MMLU-Lite86.7%자체 보고
Artificial Analysis40.0%자체 보고

Instruction Following

IFBench82.2%자체 보고

Math

AIME 202695.5%자체 보고

Reasoning

GPQANYU + Cohere + Anthropic (2023)89.5%자체 보고
ARC-AGI84.0%자체 보고
SWE-Bench Verified80.2%자체 보고
MCP Atlas79.6%자체 보고
CharXiv-R77.4%자체 보고
BrowseCompOpenAI (2025)77.4%자체 보고
Terminal-Bench 2.164.7%자체 보고
SWE-Bench ProPrinceton NLP (2024)55.9%자체 보고
SciCode48.7%자체 보고
ARC-AGI v240.1%자체 보고
Humanity's Last Exam31.6%자체 보고
CritPT8.3%자체 보고

Vision

MMMU-Pro74.0%자체 보고

AA 평가 지수

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
89.5
Lcr(Artificial Analysis)
75.7
Terminalbench V2 1
55.1
Coding Index(Artificial Analysis)
52.9
Scicode(UIUC + Argonne National Lab (2024))
49.7
Hle(Center for AI Safety + Scale AI (2025))
33.3
Intelligence Index(Artificial Analysis)
32.3
Tau Banking
18.8

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Legal
100
Finance
100
Productivity
100
Agents
100
Reasoning
100
General
100
Language
90
Multimodal
80
Search
80
Instruction Following
80
Frontend Development
80
Audio
80
Physics
70
Biology
70
Chemistry
70
Code
70
Spatial Reasoning
60
Vision
60
Math
50
Tool Calling
50

가격

입력 가격$0.3 / 1M 토큰
출력 가격$1.2 / 1M 토큰
혼합 가격 (3:1)$0.525 / 1M 토큰
캐시 읽기 가격$0.1 / 1M 토큰

속도

토큰/초117.2
첫 토큰 지연2.19s
첫 응답 지연19.25s

공급자 가격 순위

공급자 가격 순위

8개 공급자

최저가: Thinking Machines Lab최고가: LLMTR
공급자입력출력
1Thinking Machines Lab최저가
$0
$0
2Thinking Machines주요
$0.3
$1.2
3OpenRouter
$0.45
$1.2
4Kilo Gateway
$0.45
$1.2
5Baseten
$0.5
$1.2
6Vercel AI Gateway
$0.5
$1.2
7Arcee
$0.5
$1.2
8LLMTR
$0.58
$1.44

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크