Qwen2.5 Coder 32B Instruct
Summary
Qwen2.5 Coder 32B Instruct is an open-weight language model from Alibaba (Qwen), released on 11 Nov 2024. It accepts text and generates text, with a 33K-token context window and up to 29K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.06 per million input tokens and $0.20 per million output tokens (Hugging Face Inference Providers), across 3 providers we track.
Qwen coding model for software agents, repository edits, and code reasoning
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Alibaba (Qwen) |
|---|---|
| Release date | 11 Nov 2024 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | apache-2.0 [Hugging Face] |
| Parameters | 32.8B [Hugging Face] |
| Context window | 33K (32,768 tokens) [models.dev] |
| Max output | 29K tokens [models.dev] |
| Knowledge cutoff | 30 Jun 2024 [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | No [models.dev] |
| Tool calling | No [models.dev] |
| Structured output | No [models.dev] |
| Hugging Face | Qwen/Qwen2.5-Coder-32B-Instruct [Hugging Face] |
Qwen2.5 Coder 32B Instruct API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| Hugging Face Inference Providers Qwen/Qwen2.5-Coder-32B-Instruct | $0.06 | $0.20 | – | 131K | models.dev, 4 Oct 2026 |
| Cloudflare Workers AI @cf/qwen/qwen2.5-coder-32b-instruct | $0.66 | $1 | – | 33K | models.dev, 4 Oct 2026 |
| OpenRouter qwen/qwen-2.5-coder-32b-instruct | $0.66 | $1 | – | 33K | OpenRouter, 4 Oct 2026 |
Frequently asked questions
What is the context window of Qwen2.5 Coder 32B Instruct?
Qwen2.5 Coder 32B Instruct has a context window of 32,768 tokens (33K) and can generate up to 29,491 tokens in a single response, according to models.dev.
How much does the Qwen2.5 Coder 32B Instruct API cost?
Qwen2.5 Coder 32B Instruct is listed from $0.06 per million input tokens and $0.20 per million output tokens (Hugging Face Inference Providers). Prices last verified 4 Oct 2026.
Is Qwen2.5 Coder 32B Instruct open source?
Qwen2.5 Coder 32B Instruct is an open-weight model: its weights are published on Hugging Face as Qwen/Qwen2.5-Coder-32B-Instruct under the apache-2.0 license. Check the license terms before commercial use.
When was Qwen2.5 Coder 32B Instruct released?
Alibaba (Qwen) released Qwen2.5 Coder 32B Instruct on 11 Nov 2024, according to models.dev.
Which providers offer Qwen2.5 Coder 32B Instruct?
We track Qwen2.5 Coder 32B Instruct on 3 providers: Hugging Face Inference Providers, Cloudflare Workers AI, OpenRouter.
Change history
No changes detected since we started tracking this model. We re-check every source daily.
Sources
- https://openrouter.ai/qwen/qwen-2.5-coder-32b-instruct
- https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml
- https://huggingface.co/Qwen/Qwen2.5-Coder-32B-Instruct
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/Qwen/Qwen2.5-Coder-32B-Instruct.toml
- https://github.com/sst/models.dev/blob/dev/providers/cloudflare-workers-ai/models/@cf/qwen/qwen2.5-coder-32b-instruct.toml