qwen3.8-2.4t-a95b · Qwen
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a 2.4‑trillion‑parameter sparse Mixture-of-Experts (MoE) model with approximately 95 billion active parameters. It is built for autonomous, long‑duration tasks: multi‑day code runs, reproducing research papers, and self‑improvement.
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a 2.4‑trillion‑parameter sparse Mixture-of-Experts (MoE) model with approximately 95 billion active parameters. It is built for autonomous, long‑duration tasks: multi‑day code runs, reproducing research papers, and self‑improvement.
Qwen3.8 2.4t A95B has a 1,000,000 token context window.
On AIHubMix, Qwen3.8 2.4t A95B costs $2 per million input tokens and $6 per million output tokens. Cached input reads are billed at $0.5 per million tokens.
Qwen3.8 2.4t A95B accepts text input.
Qwen3.8 2.4t A95B supports tool calling, function calling, structured outputs, web search, long context and thinking. Per-protocol parameter support is listed in the capability table on this page.
Qwen3.8 2.4t A95B is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3.8-2.4t-a95b — no other code changes needed.
Qwen3.8 2.4t A95B is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8 2.4t A95B was released on August 12, 2026 by Qwen.
Qwen3.8 Omni Flash is Alibaba Cloud Qwen's next-generation native multimodal model…
Alibaba Cloud has launched the decision model decision-model-preview. This structured…
Qwen3.8-27b is an Alibaba-released dense vision-language model with open-source weights…
Qwen3.8-Max-0902 (also known as qwen3.8-max-2026-09-02) is a snapshot version of Alibaba…
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…
Use Qwen3.8 2.4t A95B via the AIHubMix unified API — one interface for every major LLM.