gme-qwen2-vl-2b-instruct · Qwen
The GME-Qwen2VL series is a unified multimodal Embedding model trained based on the Qwen2-VL multimodal large language model (MLLMs). The GME model supports three types of inputs: text, images, and image-text pairs. All these input types can generate universal vector representations and exhibit excellent retrieval performance.
The GME-Qwen2VL series is a unified multimodal Embedding model trained based on the Qwen2-VL multimodal large language model (MLLMs). The GME model supports three types of inputs: text, images, and image-text pairs. All these input types can generate universal vector representations and exhibit excellent retrieval performance.
On AIHubMix, Gme Qwen2 VL 2B Instruct costs $0.138 per million input tokens and $0.138 per million output tokens.
Gme Qwen2 VL 2B Instruct accepts text, image and video input.
Gme Qwen2 VL 2B Instruct is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gme-qwen2-vl-2b-instruct — no other code changes needed.
Gme Qwen2 VL 2B Instruct is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8 Omni Flash is Alibaba Cloud Qwen's next-generation native multimodal model…
Alibaba Cloud has launched the decision model decision-model-preview. This structured…
Qwen3.8-27b is an Alibaba-released dense vision-language model with open-source weights…
Qwen3.8-Max-0902 (also known as qwen3.8-max-2026-09-02) is a snapshot version of Alibaba…
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…
Use Gme Qwen2 VL 2B Instruct via the AIHubMix unified API — one interface for every major LLM.