Alibaba (Qwen)

Qwen3-Next 80B-A3B (Thinking)

Generally availableOpen weights131K contextReasoning

Summary

Qwen3-Next 80B-A3B (Thinking) is an open-weight language model from Alibaba (Qwen), released on Sept 2025. It accepts text and generates text, with a 131K-token context window and up to 33K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.15 per million input tokens and $1.20 per million output tokens (OpenRouter), across 5 providers we track. It is scheduled for retirement on 9 Oct 2026.

Efficient Qwen thinking model for local reasoning, math, and coding agents

Last verified 4 Oct 2026. Every figure on this page links to its source.

Specifications

DeveloperAlibaba (Qwen)
Release dateSept 2025 [models.dev]
StatusGenerally available [models.dev]
WeightsOpen weights [Hugging Face]
Licenseapache-2.0 [Hugging Face]
Parameters81.3B [Hugging Face]
Context window131K (131,072 tokens) [models.dev]
Max output33K tokens [models.dev]
Knowledge cutoffApr 2025 [models.dev]
InputText [models.dev]
OutputText [models.dev]
Reasoning modeYes [models.dev]
Tool callingYes [models.dev]
Retirement date9 Oct 2026 [OpenRouter]
Hugging FaceQwen/Qwen3-Next-80B-A3B-Thinking [Hugging Face]

Qwen3-Next 80B-A3B (Thinking) API pricing by provider

USD per million tokens. Sorted by input price.

ProviderInputOutputCached inputContextSource
Novita AI
qwen/qwen3-next-80b-a3b-thinking
$0.15$1.50–131Kmodels.dev, 4 Oct 2026
OpenRouter
qwen/qwen3-next-80b-a3b-thinking
$0.15$1.20–131KOpenRouter, 4 Oct 2026
Vercel AI Gateway
alibaba/qwen3-next-80b-a3b-thinking
$0.15$1.20–262Kmodels.dev, 4 Oct 2026
Hugging Face Inference Providers
Qwen/Qwen3-Next-80B-A3B-Thinking
$0.30$2–262Kmodels.dev, 4 Oct 2026
Alibaba Cloud Model Studio
qwen3-next-80b-a3b-thinking
$0.50$6–131Kmodels.dev, 4 Oct 2026

Frequently asked questions

What is the context window of Qwen3-Next 80B-A3B (Thinking)?

Qwen3-Next 80B-A3B (Thinking) has a context window of 131,072 tokens (131K) and can generate up to 32,768 tokens in a single response, according to models.dev.

How much does the Qwen3-Next 80B-A3B (Thinking) API cost?

Through Alibaba Cloud Model Studio, Qwen3-Next 80B-A3B (Thinking) costs $0.50 per million input tokens and $6 per million output tokens. The lowest listed price is $0.15 input / $1.20 output through OpenRouter. Prices last verified 4 Oct 2026.

Is Qwen3-Next 80B-A3B (Thinking) open source?

Qwen3-Next 80B-A3B (Thinking) is an open-weight model: its weights are published on Hugging Face as Qwen/Qwen3-Next-80B-A3B-Thinking under the apache-2.0 license. Check the license terms before commercial use.

When was Qwen3-Next 80B-A3B (Thinking) released?

Alibaba (Qwen) released Qwen3-Next 80B-A3B (Thinking) on Sept 2025, according to models.dev.

Which providers offer Qwen3-Next 80B-A3B (Thinking)?

We track Qwen3-Next 80B-A3B (Thinking) on 5 providers: Novita AI, OpenRouter, Vercel AI Gateway, Hugging Face Inference Providers, Alibaba Cloud Model Studio.

Change history

No changes detected since we started tracking this model. We re-check every source daily.

Sources