QwenQwen3 VL Flash

External Link

Qwen3 VL Flash is an Alibaba Qwen model for multimodal reasoning and tool use.

Token context
262,144 tokens
Inputs
Text, Image, Video
Outputs
Text
Released
Jan 22, 2026

Providers & Pricing

Estimate workload cost →
ProviderInput / 1MOutput / 1MCache read / 1MCache write / 1MObservedEvidence
$0.15$1.50Sep 3, 2026Provider-reported

At a Glance

Model Facts

Identity

Status
active
Developer ID
qwen3-vl-flash
Released
Jan 22, 2026
Version
Qwen3 VL Flash
License
Unknown

Capacity

Token context
262,144 tokens
Maximum output
Unknown
Total parameters
Unknown
Active parameters
Unknown
Knowledge cutoff
Not reported

Interface and Access

Inputs
Text, Image, Video
Outputs
Text
API available
Yes
Open weights
No
Self-hostable
No
Reasoning
Yes
Vision
Yes
Tool calling
Yes
Capabilities
chat, generation, reasoning, structured outputs, tools, vision

This page represents one developer model product. Dated API snapshots, serving endpoints, pricing, regions, quantizations, and service tiers attach as versioned aliases or provider details and never create duplicate public model records.

Version Lineage

PredecessorUnknown
SuccessorsUnknown
Aliasesqwen3-vl-flash, qwen3-vl-flash-2026-01-22

Explicit Unknowns

  • Maximum output
  • Architecture
  • Tokenizer
  • Knowledge cutoff

Primary Evidence

Sources and Observation Date

Comparable Peers

Related Model Comparisons

All comparisons →
APairBContext
vsfamily variantsimage, text, video
vsfamily variantsimage, text, video
vsfamily variantsimage, text, video
vscross-developer peersimage, text, video
vscross-developer peersimage, text, video
vscross-developer peersimage, text, video
vscross-developer peersimage, text, video
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peerstext
vscross-developer peerstext
vsfamily variantstext

Questions

Qwen3 VL Flash FAQs

What is Qwen3 VL Flash?+

Qwen3 VL Flash is an Alibaba Qwen model for multimodal reasoning and tool use. It is developed by Qwen and its current sourced lifecycle status is active.

What inputs and outputs does Qwen3 VL Flash support?+

Qwen3 VL Flash's current record lists Text, Image, and Video as input and Text as output.

How much context does Qwen3 VL Flash support?+

The current record lists a 262,144-token context window and does not report a maximum output length.

Is Qwen3 VL Flash available through an API?+

Qwen3 VL Flash is marked as API-available. The directly sourced provider records currently include Alibaba Cloud Model Studio.

How much does Qwen3 VL Flash cost through an API?+

The lowest directly sourced prices currently attached to Qwen3 VL Flash are $0.15 per million input tokens through Alibaba Cloud Model Studio and $1.50 per million output tokens through Alibaba Cloud Model Studio. Prices are provider-specific and should be checked against each cited observation date.

What capabilities and tasks does Qwen3 VL Flash support?+

Qwen3 VL Flash's recorded capabilities are chat, generation, reasoning, structured outputs, tools, and vision. Its supported tasks are image, reasoning, text, and video.

Can Qwen3 VL Flash be self-hosted?+

Qwen3 VL Flash is not marked as an open-weight model. Self-hosting is marked as unsupported, with no license recorded.

Does Qwen3 VL Flash have related versions or aliases?+

No predecessor is recorded. No successor is recorded. Recorded aliases are qwen3-vl-flash and qwen3-vl-flash-2026-01-22.

Which models can Qwen3 VL Flash be compared with?+

The current compatible peer set includes Qwen3 VL Plus, Qwen3.7 Flash, Qwen3.7 Max, Kimi K2.7 Code, and Seed 2.0 Mini. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons

How current is the Model Markets record for Qwen3 VL Flash?+

Qwen3 VL Flash was last verified Sep 3, 2026 from Qwen's primary source, “Qwen3 VL Flash model and pricing | Alibaba Cloud Model Studio.” The evidence is classified as official fact.

Send Feedback