Xiaomi MiMo-V2.6-Pro-RL released as 1.02T-parameter omnimodal model
Original titleXiaomiMiMo/MiMo-V2.6-Pro-RL
AISummary
Xiaomi MiMo released MiMo-V2.6-Pro-RL on Hugging Face, a sparse MoE model with 1.02T total and 42B activated parameters and a 1M-token context. The technical report says it accepts text, image, video, and audio, and was trained with a single mixed reinforcement learning run across coding, agent, visual, and cybersecurity tasks.
AIWhy it matters
The report pairs a 1.02T-parameter MoE model with an RL-based self-improvement method, useful for judging how reinforcement learning is scaled in frontier open models.
Source: Xiaomi MiMo · new models on Hugging Face · huggingface.co