# MiMo V2 Flash

> MiMo V2 Flash is Xiaomi's open-weight MoE foundation model (309B total, 15B active) for high-speed inference, coding, and agent workflows. A hybrid global/sliding-window attention layout plus multi-token prediction makes generation 2.5-3.7x faster, and it ranks at the top of the open-source field on agent and code evaluations. AIHubMix serves it behind a free quota wire.

**ID**: `xiaomi/mimo-v2-flash`  
**Creator**: Xiaomi  
**Category**: text  
**Context**: 1M tokens  
**Input**: $0.19/M  
**Output**: $0.58/M  
**Released**: 2025-12-16  
**Web page**: https://anyrouter.dev/model/xiaomi/mimo-v2-flash

**Input modalities**: text  
**Output modalities**: text  

**Tokenizer**: MiMo

**Capabilities**: chat, streaming, function-calling, reasoning, coding, agentic

**Aliases**: `xiaomi/mimo-v2-flash-free`

## Usage

**Endpoint**: `POST https://anyrouter.dev/api/v1/chat/completions`

```bash
curl https://anyrouter.dev/api/v1/chat/completions \
  -H "Authorization: Bearer $ANYROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "xiaomi/mimo-v2-flash",
  "messages": [
    {
      "role": "user",
      "content": "Say hi in 3 words."
    }
  ]
}'
```

## Providers

| Provider | Input | Output |
| --- | --- | --- |
| aihubmix-byok | $0 | $0 |
| aihubmix-byok | $0 | $0 |

## Source

- Upstream docs: https://mimo.mi.com/static/docs/news/previous-news/news20251216.md
