Xiaomi logo

MiMo V2.6 Flash

XiaomiOpen WeightsPending Human Review

MiMo V2.6 Flash is Xiaomi’s lower-compute multimodal agent tier. Its sparse model has 309B total parameters and 15B active per token, with native text, image, audio, and video understanding and a 1M context window. Reinforcement learning scales task environments and feedback across coding, general work, visual tasks, and security research. MIT-licensed weights and hosted inference support private deployment and API applications. Subsequent MOPD checkpoints are documented variants of the same named tier.

2026-09-22
309B total, 15B active
Hybrid-attention Multimodal Mixture of Experts
MIT

Specifications

Parameters
309B total, 15B active
Architecture
Hybrid-attention Multimodal Mixture of Experts
License
MIT
Context Window
1,048,576 tokens
Type
multimodal
Modalities
textimageaudiovideo

Benchmark Scores

Advanced Specifications

Model Family
MiMo
API Access
Available
Chat Interface
Available

Capabilities & Limitations

Capabilities
text generationreasoningmultimodal understanding
Known Limitations
Provider-reported evaluations need application-specific validationSparse or long-context serving can require substantial memoryCurrent information requires external retrieval
Notable Use Cases
private model deploymentdocument and code workflowslanguage-model research
Tool Use Support
Yes

Related Models