mimo-v2.6-flash

MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for greater computational efficiency. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers strong performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

Context
1.1M
Max output
131K
Input price
$0.14/1M tokens
Output price
$0.28/1M tokens

Capabilities

vision
tool call
structured output
reasoning
json mode
streaming
fine tuning
batch

Details

Model IDxiaomi/mimo-v2.6-flash
Familymimo
License—
Parameters—
Statusactive
Input modalitiestext
Output modalitiestext
Architecture—
Knowledge cutoff—
Training data cutoff—
Release date
Deprecation date—
Typechat
Reasoning tokensYes
Max input—
Open weight—
Prompt cachingSupported
Sourceofficial
Last updated

Tools

Function Callingfunction_calling
Call external functions and APIs

Pricing

Input
$0.14
↑250%$0.04
Output
$0.28
Cache write
—
Cache read
—
Batch in
—
Batch out
—
Price history · $/1M input
$0.14
$0.04
$0.14
Sep 22Oct 2Oct 9
DateInputOutputChange
Sep 22, 2026$0.14$0.28—
Oct 2, 20262b89bf1$0.04$0.28↓71%
Oct 9, 202670614e7$0.14$0.28↑250%

Price across providers

ProviderInputOutputSince launch
OpenRoutercheapest$0.14$0.28—
Vercel AI Gatewaycheapest$0.14$0.28—

Family Comparison: mimo

ModelContextMax outPricing
mimo-v2-flash262K66K$0.1$0.3
mimo-v2.51.1M131K$0.12$0.24
mimo-v2.6-flash1.1M131K$0.14$0.28
mimo-v2.5-pro1.1M131K$0.3$0.61
mimo-v2.6-pro1.1M131K$0.43$0.87
mimo-v2-pro1M131K$1$3
mimo-v2.5-pro-ultraspeed1M—$1.3$2.61
mimo-v2.6-pro-ultraspeed1M131K$4.35$8.7

API

GET/v1/models/vercel/xiaomi/mimo-v2.6-flash

mimo-v2.6-flash

xiaomi/mimo-v2.6-flash

MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for greater computational efficiency. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers strong performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

Changes · 4 entries
xiaomi/mimo-v2.6-flashupdate70614e7Oct 9, 2026, 06:39 AM
pricinginput$0.04→$0.14
xiaomi/mimo-v2.6-flashupdate2b89bf1Oct 2, 2026, 06:39 AM
pricinginput$0.14→$0.04
xiaomi/mimo-v2.6-flashupdate1a66bb0Sep 29, 2026, 06:40 AM
context_window1000000→1100000
xiaomi/mimo-v2.6-flashcreatef13f134Sep 22, 2026, 06:39 AM