Qwen3 32B
Summary
Qwen3 32B is an open-weight language model from Alibaba (Qwen), released on Apr 2025. It accepts text and generates text, with a 131K-token context window and up to 16K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.08 per million input tokens and $0.28 per million output tokens (DeepInfra), across 7 providers we track.
Dense open Qwen model for self-hosted chat, reasoning, and coding
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Alibaba (Qwen) |
|---|---|
| Release date | Apr 2025 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | apache-2.0 [Hugging Face] |
| Parameters | 32.8B [Hugging Face] |
| Context window | 131K (131,072 tokens) [models.dev] |
| Max output | 16K tokens [models.dev] |
| Knowledge cutoff | Apr 2025 [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Hugging Face | Qwen/Qwen3-32B [Hugging Face] |
Qwen3 32B API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| DeepInfra Qwen/Qwen3-32B | $0.08 | $0.28 | – | 41K | models.dev, 4 Oct 2026 |
| OpenRouter qwen/qwen3-32b | $0.08 | $0.28 | – | 41K | OpenRouter, 4 Oct 2026 |
| SiliconFlow Qwen/Qwen3-32B | $0.14 | $0.57 | – | 131K | models.dev, 4 Oct 2026 |
| Amazon Bedrock qwen.qwen3-32b-v1:0 | $0.15 | $0.60 | – | 33K | models.dev, 4 Oct 2026 |
| Vercel AI Gateway alibaba/qwen-3-32b | $0.16 | $0.64 | – | 128K | models.dev, 4 Oct 2026 |
| Hugging Face Inference Providers Qwen/Qwen3-32B | $0.29 | $0.59 | – | 131K | models.dev, 4 Oct 2026 |
| Alibaba Cloud Model Studio qwen3-32b | $0.70 | $2.80 | – | 131K | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of Qwen3 32B?
Qwen3 32B has a context window of 131,072 tokens (131K) and can generate up to 16,384 tokens in a single response, according to models.dev.
How much does the Qwen3 32B API cost?
Through Alibaba Cloud Model Studio, Qwen3 32B costs $0.70 per million input tokens and $2.80 per million output tokens. The lowest listed price is $0.08 input / $0.28 output through DeepInfra. Prices last verified 4 Oct 2026.
Is Qwen3 32B open source?
Qwen3 32B is an open-weight model: its weights are published on Hugging Face as Qwen/Qwen3-32B under the apache-2.0 license. Check the license terms before commercial use.
When was Qwen3 32B released?
Alibaba (Qwen) released Qwen3 32B on Apr 2025, according to models.dev.
Which providers offer Qwen3 32B?
We track Qwen3 32B on 7 providers: DeepInfra, OpenRouter, SiliconFlow, Amazon Bedrock, Vercel AI Gateway, Hugging Face Inference Providers, Alibaba Cloud Model Studio.
Change history
No changes detected since we started tracking this model. We re-check every source daily.
Sources
- https://openrouter.ai/qwen/qwen3-32b
- https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-32b.toml
- https://huggingface.co/Qwen/Qwen3-32B
- https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/Qwen/Qwen3-32B.toml
- https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/Qwen/Qwen3-32B.toml
- https://github.com/sst/models.dev/blob/dev/providers/amazon-bedrock/models/qwen.qwen3-32b-v1:0.toml
- https://github.com/sst/models.dev/blob/dev/providers/vercel/models/alibaba/qwen-3-32b.toml
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/Qwen/Qwen3-32B.toml