MiMo-V2-Flash (Non-reasoning)
XiaomiOpen WeightMIT · Commercial OK
Description
MiMo-V2-Flash is a powerful, efficient, and ultra-fast foundation language model that excels in reasoning, coding, and agentic scenarios. It is a Mixture-of-Experts model with 309B total parameters and 15B active parameters, featuring a hybrid attention architecture with sliding-window and full attention (5:1 ratio, 128-token window). Delivers 150 tokens/sec inference with 256k context window.
Release Date
2025-12-16
Parameters
309.0B
Context Length
262K
Modalities
text
Capability Radar
37
general
44
coding
67
reasoning
40
science
60
agents
0
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 119 | 35.0 | LS |
| Code Ranking | 220 | 51.0 | AA |
| General Ranking | 191 | 57.0 | AA |
| Science | 335 | 41.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
Tau-bench
80.3%SR
BrowseCompOpenAI (2025)
58.3%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)
38.5%SR
Terminal-Bench
30.5%SR
Biology
GPQANYU + Cohere + Anthropic (2023)
83.7%SR
Code
SWE-Bench Verified
73.4%SR
SWE-bench Multilingual
71.7%SR
Creativity
Arena-Hard v2
86.2%SR
Finance
MMLU-Pro
84.9%SR
General
LiveCodeBench v6
80.6%SR
LongBench v2
60.6%SR
MRCR
45.7%SR
Math
AIME 2025
94.1%SR
HMMT 2025
84.4%SR
Humanity's Last Exam
22.1%SR
AA Evaluation Indices
(Artificial Analysis)Math Index(Artificial Analysis)67.7
Coding Index(Artificial Analysis)49.8
Intelligence Index(Artificial Analysis)25.1
Tau2(Sierra + U Toronto + Vector Institute (2025))0.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.7
Aime 25(MAA (Mathematical Association of America))0.7
Gpqa(NYU + Cohere + Anthropic (2023))0.7
Terminalbench V2 10.6
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.4
Ifbench(Google Research (2023))0.4
Lcr(Artificial Analysis)0.3
Scicode(UIUC + Argonne National Lab (2024))0.3
Terminalbench Hard(Stanford × Laude Institute (2026))0.3
Hle(Center for AI Safety + Scale AI (2025))0.1
LLM Stats Category Scores
(LLM Stats (zeroeval))Creativity90
Writing90
Legal80
Physics80
Language80
Finance80
Healthcare80
Biology80
Chemistry80
Math70
Reasoning70
Frontend Development70
General70
Search60
Structured Output60
Tool Calling60
Long Context50
Agents50
Code50
Vision20
Pricing
Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Cache Read Price$0.0028 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
No provider data available