Identity
- Status
- active
- Developer ID
qwen3.8-flash- Released
- Not reported
- Version
- Qwen3.8-Flash
- License
- Unknown
Qwen3.8-Flash is a low-latency multimodal model with a 1M-token context window and OpenAI- and Anthropic-compatible APIs.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Cache write / 1M | Observed | Evidence |
|---|---|---|---|---|---|---|
Alibaba Cloud Model StudioAvailable | $0.80 | $2.70 | $0.10 | — | Aug 29, 2026 | Provider-reported |
OpenrouterAvailable | $0.15 | $0.47 | $0.016 | — | Sep 2, 2026 | Provider-reported |
At a Glance
qwen3.8-flashThis page represents one immutable developer model ID. Serving endpoints, pricing, regions, quantizations, and service tiers attach separately and never create duplicate model records.
Primary Evidence
Comparable Peers
| A | Pair | B | Context |
|---|---|---|---|
Qwen3.8-FlashQwen | vs | Qwen3.7-PlusQwen | family variantsimage, text, video |
Qwen3.8-FlashQwen | vs | Qwen3.8-MaxQwen | family variantsimage, text, video |
Qwen3.8-FlashQwen | vs | GLM-5V-TurboZ.ai | cross-developer peersimage, text, video |
Qwen3.8-FlashQwen | vs | Gemini 2.5 FlashGoogle DeepMind | cross-developer peerstext |
Qwen3.8-FlashQwen | vs | Gemini 3.5 FlashGoogle DeepMind | cross-developer peerstext |
Qwen3.8-FlashQwen | vs | Gemini 2.5 Flash-LiteGoogle DeepMind | cross-developer peerstext |
Qwen3.8-FlashQwen | vs | Gemini 3.1 Flash-LiteGoogle DeepMind | cross-developer peerstext |
Qwen3.8-FlashQwen | vs | Gemini 3.5 Flash-LiteGoogle DeepMind | cross-developer peerstext |
Qwen3.8-FlashQwen | vs | Gemini 3 FlashGoogle DeepMind | cross-developer peerstext |
Qwen3.8-FlashQwen | vs | Gemini 2.5 ProGoogle DeepMind | cross-developer peerstext |
Qwen3.8-FlashQwen | vs | family variantstext | |
Qwen3.8-FlashQwen | vs | Qwen3-Coder-NextQwen | family variantstext |
Questions
Qwen3.8-Flash is a low-latency multimodal model with a 1M-token context window and OpenAI- and Anthropic-compatible APIs. It is developed by Qwen and its current sourced lifecycle status is active.
Qwen3.8-Flash's current record lists Text, Image, and Video as input and Text as output.
The current record lists a 1,000,000-token context window and a maximum output of 131,072 tokens.
Qwen3.8-Flash is marked as API-available. The directly sourced provider records currently include Alibaba Cloud Model Studio and Openrouter.
The lowest directly sourced prices currently attached to Qwen3.8-Flash are $0.15 per million input tokens through Openrouter and $0.47 per million output tokens through Openrouter. Prices are provider-specific and should be checked against each cited observation date.
Qwen3.8-Flash's recorded capabilities are agents, chat, computer-use, reasoning, structured outputs, tools, and vision. Its supported tasks are coding, computer-use, image, reasoning, text, and video.
Qwen3.8-Flash is not marked as an open-weight model. Self-hosting is marked as unsupported, with no license recorded.
No predecessor is recorded. No successor is recorded. Recorded aliases are qwen3.8-flash.
The current compatible peer set includes Qwen3.7-Plus, Qwen3.8-Max, GLM-5V-Turbo, Gemini 2.5 Flash, and Gemini 3.5 Flash. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons →
Qwen3.8-Flash was last verified Aug 29, 2026 from Qwen's primary source, “qwen3.8-flash Model Info | Alibaba Cloud Model Studio.” The evidence is classified as official fact.