qwen3-vl-flash · Qwen
The Qwen3 series of compact visual-understanding models achieves an effective fusion of thinking mode and non-thinking mode, outperforming the open-source Qwen3-VL-30B-A3B with faster response speeds. It comprehensively upgrades image and video understanding, supporting ultra-long contexts such as long videos and long documents, spatial awareness, and universal object recognition; it also possesses visual 2D/3D localization capabilities and is capable of handling complex real-world tasks.
The Qwen3 series of compact visual-understanding models achieves an effective fusion of thinking mode and non-thinking mode, outperforming the open-source Qwen3-VL-30B-A3B with faster response speeds. It comprehensively upgrades image and video understanding, supporting ultra-long contexts such as long videos and long documents, spatial awareness, and universal object recognition; it also possesses visual 2D/3D localization capabilities and is capable of handling complex real-world tasks.
Qwen3 VL Flash has a 254,000 token context window.
On AIHubMix, Qwen3 VL Flash costs $0.021 per million input tokens and $0.206 per million output tokens. Cached input reads are billed at $0.0041 per million tokens.
Qwen3 VL Flash accepts text, image and video input.
Qwen3 VL Flash supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
Qwen3 VL Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3-vl-flash — no other code changes needed.
Qwen3 VL Flash is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…
Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video…
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…
Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…
Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…
Use Qwen3 VL Flash via the AIHubMix unified API — one interface for every major LLM.