# Qwen3 Coder Flash

> Qwen3 Coder Flash is a proprietary language model from Alibaba (Qwen), released on 28 Jul 2025. It accepts text and generates text, with a 1M-token context window and up to 66K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.20 per million input tokens and $0.97 per million output tokens (OpenRouter), across 2 providers we track.

Source page: https://www.aimodel.directory/models/qwen3-coder-flash
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Alibaba (Qwen)
- **Release date:** 28 Jul 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Weights:** Proprietary (API only) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Context window:** 1M (1,000,000 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Max output:** 66K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Knowledge cutoff:** Apr 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Input:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Reasoning mode:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)
- **Tool calling:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-coder-flash.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| OpenRouter | $0.20 | $0.97 | $0.04 | 1M |
| Alibaba Cloud Model Studio | $0.30 | $1.50 | – | 1M |

## FAQ

### What is the context window of Qwen3 Coder Flash?

Qwen3 Coder Flash has a context window of 1,000,000 tokens (1M) and can generate up to 65,536 tokens in a single response, according to models.dev.

### How much does the Qwen3 Coder Flash API cost?

Through Alibaba Cloud Model Studio, Qwen3 Coder Flash costs $0.30 per million input tokens and $1.50 per million output tokens. The lowest listed price is $0.20 input / $0.97 output through OpenRouter. Prices last verified 4 Oct 2026.

### Is Qwen3 Coder Flash open source?

No. Qwen3 Coder Flash is a proprietary model; Alibaba (Qwen) has not released its weights. It is available through APIs including OpenRouter, Alibaba Cloud Model Studio.

### When was Qwen3 Coder Flash released?

Alibaba (Qwen) released Qwen3 Coder Flash on 28 Jul 2025, according to models.dev.

### Which providers offer Qwen3 Coder Flash?

We track Qwen3 Coder Flash on 2 providers: OpenRouter, Alibaba Cloud Model Studio.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Qwen3 Coder Flash", https://www.aimodel.directory/models/qwen3-coder-flash