Xiaomi releases MiMo-V2.6-Flash-RL, a 309B sparse MoE model with 1M context
Original titleXiaomiMiMo/MiMo-V2.6-Flash-RL
AISummary
Xiaomi released MiMo-V2.6-Flash-RL, an efficiency-balanced checkpoint in its MiMo-V2.6 series, on Hugging Face. The model is a sparse MoE with 309B total and 15B activated parameters, supports text, image, video, and audio input, and offers a 1M-token context.
The technical report says it was trained with a single mixed reinforcement learning run across coding, agent, visual, and cybersecurity tasks.
AIWhy it matters
The report pairs its benchmark tables with the RL training method, which helps readers judge how the checkpoint's scores relate to its training approach.
Source: Xiaomi MiMo · new models on Hugging Face · huggingface.co