Skip to content

MiMo V2 Flash

Also accepted:xiaomi/mimo-v2-flash-free

MiMo V2 Flash is Xiaomi's open-weight MoE foundation model (309B total, 15B active) for high-speed inference, coding, and agent workflows. A hybrid global/sliding-window attention layout plus multi-token prediction makes generation 2.5-3.7x faster, and it ranks at the top of the open-source field on agent and code evaluations. AIHubMix serves it behind a free quota wire.

Providers
Capabilities
AIHubMix
aihubmix-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
MiMo V2 Flash share card
MiMo V2 Flash
AIHubMix upstream share card
AIHubMix upstream
Credits
Use your own key

Run MiMo V2 Flash on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length1,048,576 tokens
Max output131,000 tokens
ArchitectureTransformer
Categorytext
ReleasedDec 16, 2025
Modalities
Capabilities