mimo-v2-flash · Xiaomi
MiMo-V2-Flash is a mixture of experts (MoE) language model with a total of 309 billion parameters and 15 billion activated parameters. It is designed for high-speed inference and proxy workflows, adopting a novel hybrid attention architecture and multi-token prediction (MTP), significantly reducing inference costs while achieving state-of-the-art performance.
MiMo-V2-Flash is a mixture of experts (MoE) language model with a total of 309 billion parameters and 15 billion activated parameters. It is designed for high-speed inference and proxy workflows, adopting a novel hybrid attention architecture and multi-token prediction (MTP), significantly reducing inference costs while achieving state-of-the-art performance.
MiMo V2 Flash has a 1,048,576 token context window.
On AIHubMix, MiMo V2 Flash costs $0.1918 per million input tokens and $0.5754 per million output tokens. Cached input reads are billed at $0.0384 per million tokens.
MiMo V2 Flash accepts text, image, video and audio input.
MiMo V2 Flash supports web search. Per-protocol parameter support is listed in the capability table on this page.
MiMo V2 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to mimo-v2-flash — no other code changes needed.
MiMo V2 Flash is developed by Xiaomi. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
MiMo V2 Flash was released on December 16, 2025 by Xiaomi.
Free version: MiMo V2 Flash (free)
mimo-v2.6-flash is Xiaomi's latest model series. A fully multimodal, highly intelligent…
MiMo-V2.6-Pro is Xiaomi’s trillion-parameter, natively omni-modal flagship reasoning…
MiMo-V2.6-Pro-UltraSpeed combines the flagship performance of MiMo-V2.6-Pro and offers up…
xiaomi-mimo-v2.6-pro-free is the open free version of xiaomi-mimo-v2.6-pro. To maintain…
xiaomi-mimo-v2.6-flash-free is the open free version of xiaomi-mimo-v2.6-flash. To…
Only supports OpenAI-compatible formats.
Use MiMo V2 Flash via the AIHubMix unified API — one interface for every major LLM.