Qwen3-Next 80B-A3B (Thinking)
Summary
Qwen3-Next 80B-A3B (Thinking) is an open-weight language model from Alibaba (Qwen), released on Sept 2025. It accepts text and generates text, with a 131K-token context window and up to 33K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.15 per million input tokens and $1.20 per million output tokens (OpenRouter), across 5 providers we track. It is scheduled for retirement on 9 Oct 2026.
Efficient Qwen thinking model for local reasoning, math, and coding agents
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Alibaba (Qwen) |
|---|---|
| Release date | Sept 2025 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | apache-2.0 [Hugging Face] |
| Parameters | 81.3B [Hugging Face] |
| Context window | 131K (131,072 tokens) [models.dev] |
| Max output | 33K tokens [models.dev] |
| Knowledge cutoff | Apr 2025 [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Retirement date | 9 Oct 2026 [OpenRouter] |
| Hugging Face | Qwen/Qwen3-Next-80B-A3B-Thinking [Hugging Face] |
Qwen3-Next 80B-A3B (Thinking) API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| Novita AI qwen/qwen3-next-80b-a3b-thinking | $0.15 | $1.50 | – | 131K | models.dev, 4 Oct 2026 |
| OpenRouter qwen/qwen3-next-80b-a3b-thinking | $0.15 | $1.20 | – | 131K | OpenRouter, 4 Oct 2026 |
| Vercel AI Gateway alibaba/qwen3-next-80b-a3b-thinking | $0.15 | $1.20 | – | 262K | models.dev, 4 Oct 2026 |
| Hugging Face Inference Providers Qwen/Qwen3-Next-80B-A3B-Thinking | $0.30 | $2 | – | 262K | models.dev, 4 Oct 2026 |
| Alibaba Cloud Model Studio qwen3-next-80b-a3b-thinking | $0.50 | $6 | – | 131K | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of Qwen3-Next 80B-A3B (Thinking)?
Qwen3-Next 80B-A3B (Thinking) has a context window of 131,072 tokens (131K) and can generate up to 32,768 tokens in a single response, according to models.dev.
How much does the Qwen3-Next 80B-A3B (Thinking) API cost?
Through Alibaba Cloud Model Studio, Qwen3-Next 80B-A3B (Thinking) costs $0.50 per million input tokens and $6 per million output tokens. The lowest listed price is $0.15 input / $1.20 output through OpenRouter. Prices last verified 4 Oct 2026.
Is Qwen3-Next 80B-A3B (Thinking) open source?
Qwen3-Next 80B-A3B (Thinking) is an open-weight model: its weights are published on Hugging Face as Qwen/Qwen3-Next-80B-A3B-Thinking under the apache-2.0 license. Check the license terms before commercial use.
When was Qwen3-Next 80B-A3B (Thinking) released?
Alibaba (Qwen) released Qwen3-Next 80B-A3B (Thinking) on Sept 2025, according to models.dev.
Which providers offer Qwen3-Next 80B-A3B (Thinking)?
We track Qwen3-Next 80B-A3B (Thinking) on 5 providers: Novita AI, OpenRouter, Vercel AI Gateway, Hugging Face Inference Providers, Alibaba Cloud Model Studio.
Change history
No changes detected since we started tracking this model. We re-check every source daily.
Sources
- https://openrouter.ai/qwen/qwen3-next-80b-a3b-thinking
- https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/qwen3-next-80b-a3b-thinking.toml
- https://huggingface.co/Qwen/Qwen3-Next-80B-A3B-Thinking
- https://github.com/sst/models.dev/blob/dev/providers/novita-ai/models/qwen/qwen3-next-80b-a3b-thinking.toml
- https://github.com/sst/models.dev/blob/dev/providers/vercel/models/alibaba/qwen3-next-80b-a3b-thinking.toml
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/Qwen/Qwen3-Next-80B-A3B-Thinking.toml