Alibaba (Qwen)

Qwen3 Max Thinking

Generally available262K contextReasoning

Summary

Qwen3 Max Thinking is a language model from Alibaba (Qwen), released on 9 Feb 2026. It accepts text and generates text, with a 262K-token context window and up to 66K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.78 per million input tokens and $3.90 per million output tokens (OpenRouter), across 1 provider we track. It is scheduled for retirement on 9 Oct 2026.

Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning compute, it...

Last verified 4 Oct 2026. Every figure on this page links to its source.

Specifications

DeveloperAlibaba (Qwen)
Release date9 Feb 2026 [OpenRouter]
StatusGenerally available
Context window262K (262,144 tokens) [OpenRouter]
Max output66K tokens [OpenRouter]
InputText [OpenRouter]
OutputText [OpenRouter]
Reasoning modeYes [OpenRouter]
Tool callingYes [OpenRouter]
Structured outputYes [OpenRouter]
Retirement date9 Oct 2026 [OpenRouter]

Qwen3 Max Thinking API pricing by provider

USD per million tokens. Sorted by input price.

ProviderInputOutputCached inputContextSource
OpenRouter
qwen/qwen3-max-thinking
$0.78$3.90–262KOpenRouter, 4 Oct 2026

Frequently asked questions

What is the context window of Qwen3 Max Thinking?

Qwen3 Max Thinking has a context window of 262,144 tokens (262K) and can generate up to 65,536 tokens in a single response, according to OpenRouter.

How much does the Qwen3 Max Thinking API cost?

Qwen3 Max Thinking is listed from $0.78 per million input tokens and $3.90 per million output tokens (OpenRouter). Prices last verified 4 Oct 2026.

When was Qwen3 Max Thinking released?

Alibaba (Qwen) released Qwen3 Max Thinking on 9 Feb 2026, according to OpenRouter.

Change history

No changes detected since we started tracking this model. We re-check every source daily.

Sources