Qwen 3.8 Flash
High-volume agentic workloads that need Qwen3.8 quality at Flash economics.
What are Qwen 3.8 Flash's specs and price?
Qwen 3.8 Flash, built by Qwen, ships a 984K-token context window and a 131K-token max output, released 2026-09. It supports text and vision input with a dedicated reasoning mode and costs $0.23 per million blended tokens, the 8th-cheapest of 49 models we track.
What are Qwen 3.8 Flash's specs?
| Context window | 984K tokens |
| Max output | 131K tokens |
| Modalities | text, vision |
| Extended thinking | Yes |
| Released | 2026-09 |
| Knowledge cutoff | Not published |
| Provider | Qwen |
Verified 2026-10-01 — source.
Where does Qwen 3.8 Flash rank?
What are Qwen 3.8 Flash's strengths?
- Low-latency Qwen3.8 tier at $0.15/$0.47 per million tokens
- Reasoning enabled by default
- Native text-and-image input
What else should you know about Qwen 3.8 Flash?
What are common questions about Qwen 3.8 Flash?
What is Qwen 3.8 Flash's context window?
Qwen 3.8 Flash has a 984K-token context window and a 131K-token max output — the 24th-largest context of the 49 current models we track. Source: https://www.alibabacloud.com/help/en/model-studio/models, verified 2026-10-01.
Does Qwen 3.8 Flash support vision or audio input?
Yes — Qwen 3.8 Flash accepts vision input in addition to text.
Does Qwen 3.8 Flash have a reasoning or extended-thinking mode?
Yes — Qwen 3.8 Flash exposes a dedicated reasoning mode for multi-step problems.
When was Qwen 3.8 Flash released, and what is its knowledge cutoff?
Qwen 3.8 Flash was released 2026-09.
How much does Qwen 3.8 Flash cost, and who provides it?
Qwen 3.8 Flash is served by Qwen at $0.23/M blended tokens (3:1 input:output) — the 8th-cheapest of 49 current models. Full pricing breakdown: /llm-api-pricing/qwen3-8-flash.
