MiMo V2 Flash is Xiaomi's open-weight MoE foundation model (309B total, 15B active) for high-speed inference, coding, and agent workflows. A hybrid global/sliding-window attention layout plus multi-token prediction makes generation 2.5-3.7x faster, and it ranks at the top of the open-source field on agent and code evaluations. AIHubMix serves it behind a free quota wire.
Share cards2 images

