MiMo-V2-Omni
XiaomiProprietary
描述
MiMo-V2-Omni is Xiaomi's omni foundation model uniting frontier multimodal understanding with strong agentic capability. It fuses dedicated image, video, and audio encoders into a single shared backbone, processing all modalities simultaneously. Natively supports structured tool calling, function execution, and UI grounding. Supports over 10 hours of continuous audio understanding and 256K token context window.
发布日期
2026-03-19
参数规模
—
上下文长度
262K
支持模态
audio, image, pdf, text, video
能力雷达图
24
general
70
coding
83
reasoning
64
science
70
agents
85
multimodal
排行榜排名
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Agents
MM-BrowserComp
52.0%自报
OmniGAIA
49.8%自报
Code
PinchBench
81.2%自报
Claw-Eval
54.8%自报
Reasoning
SWE-Bench Verified
74.8%自报
AA 评测指数
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))91.2
Gpqa(NYU + Cohere + Anthropic (2023))82.8
Lcr(Artificial Analysis)75.0
Ifbench(Google Research (2023))53.5
Terminalbench Hard(Stanford × Laude Institute (2026))34.8
Intelligence Index(Artificial Analysis)23.9
Hle(Center for AI Safety + Scale AI (2025))22.1
LLM Stats 分类评分
(LLM Stats (zeroeval))Reasoning70
Frontend Development70
General70
Agents70
Code70
定价
输入价格免费
输出价格免费
混合价格(3:1)免费
缓存读取价格$0.0028 / 1M tokens
速度
Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s
供应商价格排行
供应商价格排行
1 个供应商
供应商输入输出
1Xiaomi
$0.14
$0.28
比较该模型在不同 API 供应商之间的定价。