qwen3-vl-30b-a3b-thinking · Qwen
The Qwen3-VL series’ second-largest MoE model Thinking version offers fast response speed, stronger multimodal understanding and reasoning, visual agent capabilities, and ultra-long context support for long videos and long documents; it features comprehensive upgrades in image/video understanding, spatial perception, and universal recognition abilities, making it capable of handling complex real-world tasks.
The Qwen3-VL series’ second-largest MoE model Thinking version offers fast response speed, stronger multimodal understanding and reasoning, visual agent capabilities, and ultra-long context support for long videos and long documents; it features comprehensive upgrades in image/video understanding, spatial perception, and universal recognition abilities, making it capable of handling complex real-world tasks.
Qwen3 VL 30B A3B Thinking has a 128,000 token context window.
On AIHubMix, Qwen3 VL 30B A3B Thinking costs $0.103 per million input tokens and $1.028 per million output tokens.
Qwen3 VL 30B A3B Thinking accepts text, image and video input.
Qwen3 VL 30B A3B Thinking supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
Qwen3 VL 30B A3B Thinking is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3-vl-30b-a3b-thinking — no other code changes needed.
Qwen3 VL 30B A3B Thinking is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…
Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…
Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…
Qwen 3.8 Max(qwen3.8-max) is Alibaba Cloud’s flagship native vision-language model, built…
Qwen 3.8 Max Preview(Qwen3.8-Max-Preview) is the latest-generation foundation model in…
qwen-audio-3.0-tts-flash is a high-performance speech synthesis large model optimized for…
Use Qwen3 VL 30B A3B Thinking via the AIHubMix unified API — one interface for every major LLM.