Qwen2.5 32B Instruct
Summary
Qwen2.5 32B Instruct is an open-weight language model from Alibaba (Qwen), released on Sept 2024. It accepts text and generates text, with a 131K-token context window and up to 8K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.70 per million input tokens and $2.80 per million output tokens (Alibaba Cloud Model Studio), across 1 provider we track.
Qwen instruction model for multilingual chat, reasoning, and tool use
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Alibaba (Qwen) |
|---|---|
| Release date | Sept 2024 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [models.dev] |
| Context window | 131K (131,072 tokens) [models.dev] |
| Max output | 8K tokens [models.dev] |
| Knowledge cutoff | Apr 2024 [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | No [models.dev] |
| Tool calling | Yes [models.dev] |
Qwen2.5 32B Instruct API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| Alibaba Cloud Model Studio qwen2-5-32b-instruct | $0.70 | $2.80 | – | 131K | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of Qwen2.5 32B Instruct?
Qwen2.5 32B Instruct has a context window of 131,072 tokens (131K) and can generate up to 8,192 tokens in a single response, according to models.dev.
How much does the Qwen2.5 32B Instruct API cost?
Through Alibaba Cloud Model Studio, Qwen2.5 32B Instruct costs $0.70 per million input tokens and $2.80 per million output tokens. Prices last verified 4 Oct 2026.
Is Qwen2.5 32B Instruct open source?
Qwen2.5 32B Instruct is an open-weight model. Check the license terms before commercial use.
When was Qwen2.5 32B Instruct released?
Alibaba (Qwen) released Qwen2.5 32B Instruct on Sept 2024, according to models.dev.
Change history
No changes detected since we started tracking this model. We re-check every source daily.