MiMo-V2.6-Flash-RL
Xiaomi’s omnimodal 309B MoE model for coding and agents, with 15B active parameters, a 1M-token context window, and reinforcement learning.
Model details
View repositoryMiMo-V2.6-Flash-RL is Xiaomi’s efficiency-balanced omnimodal AI model, built on a sparse Mixture-of-Experts architecture with 309B total parameters (15B activated) and a hybrid sliding-window/global attention design supporting context windows up to 1M tokens. It natively processes text, images, video, and audio through dedicated vision and audio encoders while producing text outputs. Trained through large-scale mixed reinforcement learning, it excels at multimodal perception, long-context reasoning, coding, cybersecurity, and agentic workflows involving multi-step tool use.