Qwen/QwQ-32B

Qwen

Silicon-based flow provision

API Pricing

Input$0.14 / 1M tokens
Output$0.56 / 1M tokens

Specifications

Featurestool calling, function calling, structured outputs

Frequently asked questions

What is Qwen/QwQ-32B?

Silicon-based flow provision

How much does Qwen/QwQ-32B cost?

On AIHubMix, Qwen/QwQ-32B costs $0.14 per million input tokens and $0.56 per million output tokens.

What capabilities does Qwen/QwQ-32B support?

Qwen/QwQ-32B supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call Qwen/QwQ-32B via API?

Qwen/QwQ-32B is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to Qwen/QwQ-32B — no other code changes needed.

Who created Qwen/QwQ-32B?

Qwen/QwQ-32B is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Qwen

See all Qwen models →

Qwen3.8 Omni Flash

by Qwen

Qwen3.8 Omni Flash is Alibaba Cloud Qwen's next-generation native multimodal model…

$0.1126/1M in · $0.38/1M out
1,000,000 tokens context

Decision Model Preview

by Qwen

Alibaba Cloud has launched the decision model decision-model-preview. This structured…

64,000 tokens context

Qwen 3.8 27B

by Qwen

Qwen3.8-27b is an Alibaba-released dense vision-language model with open-source weights…

$1.1/1M in · $1.65/1M out
131,072 tokens context

Qwen3.8 Max 2026 09-02

by Qwen

Qwen3.8-Max-0902 (also known as qwen3.8-max-2026-09-02) is a snapshot version of Alibaba…

$1.69/1M in · $5.07/1M out
1,000,000 tokens context

Qwen3.8 Flash

by Qwen

Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…

$0.1126/1M in · $0.38/1M out
1,000,000 tokens context

Wan3.0 Video

by Qwen

Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…

$2/1M in · $2/1M out

Use Qwen/QwQ-32B via the AIHubMix unified API — one interface for every major LLM.