Qwen3 8B

qwen3-8b · Qwen

Achieves effective integration of thinking and non-thinking modes, enabling mode switching during conversations. Its reasoning ability reaches state-of-the-art (SOTA) levels among models of the same scale, and its general capability significantly surpasses Qwen2.5-7B.

API Pricing

Input$0.08 / 1M tokens
Output$0.8 / 1M tokens

Frequently asked questions

What is Qwen3 8B?

Achieves effective integration of thinking and non-thinking modes, enabling mode switching during conversations. Its reasoning ability reaches state-of-the-art (SOTA) levels among models of the same scale, and its general capability significantly surpasses Qwen2.5-7B.

How much does Qwen3 8B cost?

On AIHubMix, Qwen3 8B costs $0.08 per million input tokens and $0.8 per million output tokens.

How do I call Qwen3 8B via API?

Qwen3 8B is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3-8b — no other code changes needed.

Who created Qwen3 8B?

Qwen3 8B is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Qwen

See all Qwen models →

Qwen3.8 2.4t A95B

by Qwen

Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…

$2/1M in · $6/1M out
262,000 tokens context

Qwen Image 3.0

by Qwen

Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…

$2/1M in

Qwen Image 3.0 Pro

by Qwen

Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…

$2/1M in

Qwen3.8 Max

by Qwen

Qwen 3.8 Max(qwen3.8-max) is Alibaba Cloud’s flagship native vision-language model, built…

$1.69/1M in · $5.07/1M out
991,000 tokens context

Qwen3.8 Max Preview

by Qwen

Qwen 3.8 Max Preview(Qwen3.8-Max-Preview) is the latest-generation foundation model in…

$0.338/1M in · $1.014/1M out
983,616 tokens context

Qwen Audio 3.0 Tts Flash

by Qwen

qwen-audio-3.0-tts-flash is a high-performance speech synthesis large model optimized for…

$14.2/1M in · $14.2/1M out

Use Qwen3 8B via the AIHubMix unified API — one interface for every major LLM.