# Qwen2.5 Coder 32B Instruct

> Qwen2.5 Coder 32B Instruct is an open-weight language model from Alibaba (Qwen), released on 11 Nov 2024. It accepts text and generates text, with a 33K-token context window and up to 29K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.06 per million input tokens and $0.20 per million output tokens (Hugging Face Inference Providers), across 3 providers we track.

Source page: https://www.aimodel.directory/models/qwen2-5-coder-32b-instruct
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Alibaba (Qwen)
- **Release date:** 11 Nov 2024 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Weights:** Open weights (source: Hugging Face, https://huggingface.co/Qwen/Qwen2.5-Coder-32B-Instruct)
- **License:** apache-2.0 (source: Hugging Face, https://huggingface.co/Qwen/Qwen2.5-Coder-32B-Instruct)
- **Parameters:** 32.8B (source: Hugging Face, https://huggingface.co/Qwen/Qwen2.5-Coder-32B-Instruct)
- **Context window:** 33K (32,768 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Max output:** 29K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Knowledge cutoff:** 30 Jun 2024 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Input:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Reasoning mode:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Tool calling:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Structured output:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/openrouter/models/qwen/qwen-2.5-coder-32b-instruct.toml)
- **Hugging Face:** Qwen/Qwen2.5-Coder-32B-Instruct (source: Hugging Face, https://huggingface.co/Qwen/Qwen2.5-Coder-32B-Instruct)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| Hugging Face Inference Providers | $0.06 | $0.20 | – | 131K |
| Cloudflare Workers AI | $0.66 | $1 | – | 33K |
| OpenRouter | $0.66 | $1 | – | 33K |

## FAQ

### What is the context window of Qwen2.5 Coder 32B Instruct?

Qwen2.5 Coder 32B Instruct has a context window of 32,768 tokens (33K) and can generate up to 29,491 tokens in a single response, according to models.dev.

### How much does the Qwen2.5 Coder 32B Instruct API cost?

Qwen2.5 Coder 32B Instruct is listed from $0.06 per million input tokens and $0.20 per million output tokens (Hugging Face Inference Providers). Prices last verified 4 Oct 2026.

### Is Qwen2.5 Coder 32B Instruct open source?

Qwen2.5 Coder 32B Instruct is an open-weight model: its weights are published on Hugging Face as Qwen/Qwen2.5-Coder-32B-Instruct under the apache-2.0 license. Check the license terms before commercial use.

### When was Qwen2.5 Coder 32B Instruct released?

Alibaba (Qwen) released Qwen2.5 Coder 32B Instruct on 11 Nov 2024, according to models.dev.

### Which providers offer Qwen2.5 Coder 32B Instruct?

We track Qwen2.5 Coder 32B Instruct on 3 providers: Hugging Face Inference Providers, Cloudflare Workers AI, OpenRouter.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Qwen2.5 Coder 32B Instruct", https://www.aimodel.directory/models/qwen2-5-coder-32b-instruct