MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for greater computational efficiency. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers strong performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.
function_calling| Provider | Input | Output | Since launch |
|---|---|---|---|
| OpenRoutercheapest | $0.14 | $0.28 | — |
| Vercel AI Gatewaycheapest | $0.14 | $0.28 | — |
| Model | Context | Max out | Pricing |
|---|---|---|---|
| mimo-v2-flash | 262K | 66K | $0.1$0.3 |
| mimo-v2.5 | 1.1M | 131K | $0.12$0.24 |
| mimo-v2.6-flash | 1.1M | 131K | $0.14$0.28 |
| mimo-v2.5-pro | 1.1M | 131K | $0.3$0.61 |
| mimo-v2.6-pro | 1.1M | 131K | $0.43$0.87 |
| mimo-v2-pro | 1M | 131K | $1$3 |
| mimo-v2.5-pro-ultraspeed | 1M | — | $1.3$2.61 |
| mimo-v2.6-pro-ultraspeed | 1M | 131K | $4.35$8.7 |
/v1/models/vercel/xiaomi/mimo-v2.6-flashMiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for greater computational efficiency. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers strong performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.