qwen3.8-omni-flash · Qwen
Qwen3.8 Omni Flash is Alibaba Cloud Qwen's next-generation native multimodal model, supporting 1M-length sequences and able to directly accept text, images, audio, and video as input. It is based on the Qwen3.8-Flash-Next architecture. The model is designed for agent capabilities in real productivity scenarios: while offering agentic abilities such as programming, text knowledge work, and GUI operation, it also achieves significant results in audio/video-centered agentic applications—such as video editing, music video creation, film production and narration, audio/video-to-text-and-image summarization, and audio/video dialogue—that require integrated processing of text, images, audio, and video. It supports 2-channel and 4-channel spatial audio parsing, is compatible with DashScope and OpenAI protocols, and it is recommended to install the accompanying Qwen-MM-Plugins to facilitate agent frameworks' access to native multimodal capabilities.
Qwen3.8 Omni Flash is Alibaba Cloud Qwen's next-generation native multimodal model, supporting 1M-length sequences and able to directly accept text, images, audio, and video as input. It is based on the Qwen3.8-Flash-Next architecture. The model is designed for agent capabilities in real productivity scenarios: while offering agentic abilities such as programming, text knowledge work, and GUI operation, it also achieves significant results in audio/video-centered agentic applications—such as video editing, music video creation, film production and narration, audio/video-to-text-and-image summarization, and audio/video dialogue—that require integrated processing of text, images, audio, and video. It supports 2-channel and 4-channel spatial audio parsing, is compatible with DashScope and OpenAI protocols, and it is recommended to install the accompanying Qwen-MM-Plugins to facilitate agent frameworks' access to native multimodal capabilities.
Qwen3.8 Omni Flash has a 1,000,000 token context window.
On AIHubMix, Qwen3.8 Omni Flash costs $0.1126 per million input tokens and $0.38 per million output tokens. Cached input reads are billed at $0.0141 per million tokens.
Qwen3.8 Omni Flash accepts text, image, audio and video input.
Qwen3.8 Omni Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3.8-omni-flash — no other code changes needed.
Qwen3.8 Omni Flash is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8 Omni Flash was released on August 26, 2026 by Qwen.
Alibaba Cloud has launched the decision model decision-model-preview. This structured…
Qwen3.8-27b is an Alibaba-released dense vision-language model with open-source weights…
Qwen3.8-Max-0902 (also known as qwen3.8-max-2026-09-02) is a snapshot version of Alibaba…
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…
Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video…
Use Qwen3.8 Omni Flash via the AIHubMix unified API — one interface for every major LLM.