Qwen3.8 Flash Next
Summary
Qwen3.8 Flash Next is an open-weight language model from Alibaba (Qwen), released on 27 Aug 2026. It accepts text, images and video and generates text, with a 262K-token context window and up to 131K output tokens per response.
Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Alibaba (Qwen) |
|---|---|
| Release date | 27 Aug 2026 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [models.dev] |
| Context window | 262K (262,144 tokens) [models.dev] |
| Max output | 131K tokens [models.dev] |
| Input | Text, images and video [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Structured output | Yes [models.dev] |
Qwen3.8 Flash Next API pricing by provider
USD per million tokens. Sorted by input price.
We do not currently track a paid API listing for this model.
Frequently asked questions
What is the context window of Qwen3.8 Flash Next?
Qwen3.8 Flash Next has a context window of 262,144 tokens (262K) and can generate up to 131,072 tokens in a single response, according to models.dev.
Is Qwen3.8 Flash Next open source?
Qwen3.8 Flash Next is an open-weight model. Check the license terms before commercial use.
When was Qwen3.8 Flash Next released?
Alibaba (Qwen) released Qwen3.8 Flash Next on 27 Aug 2026, according to models.dev.
Change history
No changes detected since we started tracking this model. We re-check every source daily.