Qwen3 VL Flash 2026 01-22

qwen3-vl-flash-2026-01-22 · Qwen

The Qwen3 series of compact visual-understanding models achieves an effective fusion of thinking mode and non-thinking mode, outperforming the open-source Qwen3-VL-30B-A3B with faster response speeds. It comprehensively upgrades image and video understanding, supporting ultra-long contexts such as long videos and long documents, spatial awareness, and universal object recognition; it also possesses visual 2D/3D localization capabilities and is capable of handling complex real-world tasks.

API Pricing

Input$0.021 / 1M tokens
Output$0.206 / 1M tokens

Specifications

Context254K tokens
Modalitiestext, image, video
Featurestool calling, function calling, structured outputs

Frequently asked questions

What is Qwen3 VL Flash 2026 01-22?

The Qwen3 series of compact visual-understanding models achieves an effective fusion of thinking mode and non-thinking mode, outperforming the open-source Qwen3-VL-30B-A3B with faster response speeds. It comprehensively upgrades image and video understanding, supporting ultra-long contexts such as long videos and long documents, spatial awareness, and universal object recognition; it also possesses visual 2D/3D localization capabilities and is capable of handling complex real-world tasks.

What is the context length of Qwen3 VL Flash 2026 01-22?

Qwen3 VL Flash 2026 01-22 has a 254,000 token context window.

How much does Qwen3 VL Flash 2026 01-22 cost?

On AIHubMix, Qwen3 VL Flash 2026 01-22 costs $0.021 per million input tokens and $0.206 per million output tokens.

What modalities does Qwen3 VL Flash 2026 01-22 support?

Qwen3 VL Flash 2026 01-22 accepts text, image and video input.

What capabilities does Qwen3 VL Flash 2026 01-22 support?

Qwen3 VL Flash 2026 01-22 supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call Qwen3 VL Flash 2026 01-22 via API?

Qwen3 VL Flash 2026 01-22 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3-vl-flash-2026-01-22 — no other code changes needed.

Who created Qwen3 VL Flash 2026 01-22?

Qwen3 VL Flash 2026 01-22 is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Qwen

See all Qwen models →

Qwen3.8 Flash

by Qwen

Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…

$0.113/1M in · $0.38/1M out
1,000,000 tokens context

Wan3.0 Video

by Qwen

Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…

$2/1M in · $2/1M out

Wan3.0 Video Prime

by Qwen

Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video…

$2/1M in · $2/1M out

Qwen3.8 2.4t A95B

by Qwen

Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…

$2/1M in · $6/1M out
262,000 tokens context

Qwen Image 3.0

by Qwen

Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…

$2/1M in

Qwen Image 3.0 Pro

by Qwen

Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…

$2/1M in

Use Qwen3 VL Flash 2026 01-22 via the AIHubMix unified API — one interface for every major LLM.