Xiaomi releases open-source MiMo-V2-Flash MoE model for reasoning and coding
Original titleXiaomi MiMo-V2-Flash
AISummary
Xiaomi released and open-sourced MiMo-V2-Flash, a Mixture-of-Experts model with 309B total and 15B active parameters, under the MIT license.
The company reports 73.4% on SWE-Bench Verified, the top score among open-source models, and inference at 150 tokens per second for $0.1 per million input tokens and $0.3 per million output tokens. It supports a hybrid thinking mode and a 256k context window.
AIWhy it matters
The post gives architecture, speculative decoding speedup, and pricing figures, which help readers judge how the efficiency claims are achieved and what they cost.
Source: Xiaomi MiMo · mimo.xiaomi.comPublished · added here