X

Xiaomi: mimo-v2.6-flash

mimo-v2.6-flash
1.0M context131.1K outReasoningToolsVisionAudio inVideoCacheStructured
Released Sep 21, 2026Updated Sep 22, 2026

MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for greater computational efficiency. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers strong performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

Mode chatTokenizer OtherQuantization fp8XiaomiMiMo/MiMo-V2.6-Flash-RL

Pricing

Input price
$0.28/ 1M tokens
Output price
$0.56/ 1M tokens
Compatible endpoints openaiVendor Xiaomi

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providers0
include_reasoningAll providers-
logit_biasAll providers-
max_tokensAll providers-
min_pAll providers-
presence_penaltyAll providers-
reasoningAll providers-
repetition_penaltyAll providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_pAll providers0.95

Frequently asked questions

Similar models