X

Xiaomi: mimo-v2-omni:free

mimo-v2-omni:free
262.1K contextReasoningToolsVisionCacheWeb search
Released Mar 18, 2026Updated Sep 4, 2026

MiMo-V2-Omni is a frontier omni-modal model that natively processes image, video, and audio inputs within a unified architecture. It combines strong multimodal perception with agentic capability - visual grounding, multi-step planning, tool use, and code execution - making it well-suited for complex real-world tasks that span modalities, 256K context window.

Mode chat

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor Xiaomi

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Frequently asked questions

Similar models