MiMo V2.6 Flash
MiMo V2.6 Flash is Xiaomi’s lower-compute multimodal agent tier. Its sparse model has 309B total parameters and 15B active per token, with native text, image, audio, and video understanding and a 1M context window. Reinforcement learning scales task environments and feedback across coding, general work, visual tasks, and security research. MIT-licensed weights and hosted inference support private deployment and API applications. Subsequent MOPD checkpoints are documented variants of the same named tier.
2026-09-22
309B total, 15B active
Hybrid-attention Multimodal Mixture of Experts
MIT
Specifications
- Parameters
- 309B total, 15B active
- Architecture
- Hybrid-attention Multimodal Mixture of Experts
- License
- MIT
- Context Window
- 1,048,576 tokens
- Type
- multimodal
- Modalities
- textimageaudiovideo
Benchmark Scores
Advanced Specifications
- Model Family
- MiMo
- API Access
- Available
- Chat Interface
- Available
Capabilities & Limitations
- Capabilities
- text generationreasoningmultimodal understanding
- Known Limitations
- Provider-reported evaluations need application-specific validationSparse or long-context serving can require substantial memoryCurrent information requires external retrieval
- Notable Use Cases
- private model deploymentdocument and code workflowslanguage-model research
- Tool Use Support
- Yes