Qwen3.8 Max

qwen3.8-max · Qwen

Qwen 3.8 Max(qwen3.8-max) is Alibaba Cloud’s flagship native vision-language model, built on a 2.4-trillion-parameter Mixture-of-Experts (MoE) architecture and supporting context windows of up to 1 million tokens. It is well suited for complex multimodal understanding, advanced reasoning, software development, agentic workflows, and long-context processing. At a similar price to Qwen3.7-Max, Qwen3.8-Max delivers significant improvements in reasoning, coding, and agent capabilities, with overall performance comparable to today’s leading models.

API Pricing

Input$1.69 / 1M tokens
Output$5.07 / 1M tokens
Cache read$0.169 / 1M tokens

Specifications

Context984K tokens
Max output131K tokens
Modalitiestext, image, video
CapabilitiesThinking, Streaming, Tool calling, Web search, Code interpreter, Structured outputs, Prompt caching

Frequently asked questions

What is Qwen3.8 Max?

Qwen 3.8 Max(qwen3.8-max) is Alibaba Cloud’s flagship native vision-language model, built on a 2.4-trillion-parameter Mixture-of-Experts (MoE) architecture and supporting context windows of up to 1 million tokens. It is well suited for complex multimodal understanding, advanced reasoning, software development, agentic workflows, and long-context processing. At a similar price to Qwen3.7-Max, Qwen3.8-Max delivers significant improvements in reasoning, coding, and agent capabilities, with overall performance comparable to today’s leading models.

What is the context length of Qwen3.8 Max?

Qwen3.8 Max has a 983,616 token context window. It supports up to 131,072 output tokens.

How much does Qwen3.8 Max cost?

On AIHubMix, Qwen3.8 Max costs $1.69 per million input tokens and $5.07 per million output tokens. Cached input reads are billed at $0.169 per million tokens.

What modalities does Qwen3.8 Max support?

Qwen3.8 Max accepts text, image and video input.

What capabilities does Qwen3.8 Max support?

Qwen3.8 Max supports tool calling, function calling, structured outputs, web search, long context and thinking. Per-protocol parameter support is listed in the capability table on this page.

How do I call Qwen3.8 Max via API?

Qwen3.8 Max is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3.8-max — no other code changes needed.

Who created Qwen3.8 Max?

Qwen3.8 Max is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was Qwen3.8 Max released?

Qwen3.8 Max was released on August 3, 2026 by Qwen.

More models from Qwen

See all Qwen models →

Qwen3.8 2.4t A95B

by Qwen

Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…

$2/1M in · $6/1M out
262,000 tokens context

Qwen Image 3.0

by Qwen

Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…

$2/1M in

Qwen Image 3.0 Pro

by Qwen

Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…

$2/1M in

Qwen3.8 Max Preview

by Qwen

Qwen 3.8 Max Preview(Qwen3.8-Max-Preview) is the latest-generation foundation model in…

$0.338/1M in · $1.014/1M out
983,616 tokens context

Qwen Audio 3.0 Tts Flash

by Qwen

qwen-audio-3.0-tts-flash is a high-performance speech synthesis large model optimized for…

$14.2/1M in · $14.2/1M out

Qwen Audio 3.0 Tts Plus

by Qwen

qwen-audio-3.0-tts-plus is a high-performance speech synthesis large model designed for…

$15/1M in · $15/1M out

Use Qwen3.8 Max via the AIHubMix unified API — one interface for every major LLM.