Identity
- Status
- active
- Developer ID
qwen3-vl-flash- Released
- Jan 22, 2026
- Version
- Qwen3 VL Flash
- License
- Unknown
Qwen3 VL Flash is an Alibaba Qwen model for multimodal reasoning and tool use.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Cache write / 1M | Observed | Evidence |
|---|---|---|---|---|---|---|
Alibaba Cloud Model StudioAvailable | $0.15 | $1.50 | — | — | Sep 3, 2026 | Provider-reported |
At a Glance
qwen3-vl-flashThis page represents one developer model product. Dated API snapshots, serving endpoints, pricing, regions, quantizations, and service tiers attach as versioned aliases or provider details and never create duplicate public model records.
Primary Evidence
Comparable Peers
| A | Pair | B | Context |
|---|---|---|---|
Qwen3 VL FlashQwen | vs | Qwen3 VL PlusQwen | family variantsimage, text, video |
Qwen3 VL FlashQwen | vs | Qwen3.7 FlashQwen | family variantsimage, text, video |
Qwen3 VL FlashQwen | vs | Qwen3.7 MaxQwen | family variantsimage, text, video |
Qwen3 VL FlashQwen | vs | Kimi K2.7 CodeMoonshot AI | cross-developer peersimage, text, video |
Qwen3 VL FlashQwen | vs | Seed 2.0 MiniByteDance Seed | cross-developer peersimage, text, video |
Qwen3 VL FlashQwen | vs | Seed 2.1 EvolvingByteDance Seed | cross-developer peersimage, text, video |
Qwen3 VL FlashQwen | vs | Seed 2.1 TurboByteDance Seed | cross-developer peersimage, text, video |
Qwen3 VL FlashQwen | vs | Claude Sonnet 4.5Anthropic | cross-developer peersimage, text |
Qwen3 VL FlashQwen | vs | o3OpenAI | cross-developer peersimage, text |
Qwen3 VL FlashQwen | vs | Gemini 2.5 FlashGoogle DeepMind | cross-developer peerstext |
Qwen3 VL FlashQwen | vs | Gemini 3.5 Flash-LiteGoogle DeepMind | cross-developer peerstext |
Qwen3 VL FlashQwen | vs | family variantstext |
Questions
Qwen3 VL Flash is an Alibaba Qwen model for multimodal reasoning and tool use. It is developed by Qwen and its current sourced lifecycle status is active.
Qwen3 VL Flash's current record lists Text, Image, and Video as input and Text as output.
The current record lists a 262,144-token context window and does not report a maximum output length.
Qwen3 VL Flash is marked as API-available. The directly sourced provider records currently include Alibaba Cloud Model Studio.
The lowest directly sourced prices currently attached to Qwen3 VL Flash are $0.15 per million input tokens through Alibaba Cloud Model Studio and $1.50 per million output tokens through Alibaba Cloud Model Studio. Prices are provider-specific and should be checked against each cited observation date.
Qwen3 VL Flash's recorded capabilities are chat, generation, reasoning, structured outputs, tools, and vision. Its supported tasks are image, reasoning, text, and video.
Qwen3 VL Flash is not marked as an open-weight model. Self-hosting is marked as unsupported, with no license recorded.
No predecessor is recorded. No successor is recorded. Recorded aliases are qwen3-vl-flash and qwen3-vl-flash-2026-01-22.
The current compatible peer set includes Qwen3 VL Plus, Qwen3.7 Flash, Qwen3.7 Max, Kimi K2.7 Code, and Seed 2.0 Mini. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons →
Qwen3 VL Flash was last verified Sep 3, 2026 from Qwen's primary source, “Qwen3 VL Flash model and pricing | Alibaba Cloud Model Studio.” The evidence is classified as official fact.