gme-qwen2-vl-2b-instruct · Qwen
The GME-Qwen2VL series is a unified multimodal Embedding model trained based on the Qwen2-VL multimodal large language model (MLLMs). The GME model supports three types of inputs: text, images, and image-text pairs. All these input types can generate universal vector representations and exhibit excellent retrieval performance.
The GME-Qwen2VL series is a unified multimodal Embedding model trained based on the Qwen2-VL multimodal large language model (MLLMs). The GME model supports three types of inputs: text, images, and image-text pairs. All these input types can generate universal vector representations and exhibit excellent retrieval performance.
On AIHubMix, Gme Qwen2 VL 2B Instruct costs $0.138 per million input tokens and $0.138 per million output tokens.
Gme Qwen2 VL 2B Instruct accepts text, image and video input.
Gme Qwen2 VL 2B Instruct is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gme-qwen2-vl-2b-instruct — no other code changes needed.
Gme Qwen2 VL 2B Instruct is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…
Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video…
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…
Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…
Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…
Use Gme Qwen2 VL 2B Instruct via the AIHubMix unified API — one interface for every major LLM.