Qwen3.8 2.4T A95B
Summary
Qwen3.8 2.4T A95B is an open-weight language model from Alibaba (Qwen), released on 12 Aug 2026. It accepts text and generates text, with a 262K-token context window and up to 131K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $2 per million input tokens and $6 per million output tokens (DeepInfra), across 6 providers we track.
Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Alibaba (Qwen) |
|---|---|
| Release date | 12 Aug 2026 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | qwen3.8-max [Hugging Face] |
| Parameters | 2.4T [Hugging Face] |
| Context window | 262K (262,144 tokens) [models.dev] |
| Max output | 131K tokens [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Structured output | Yes [models.dev] |
| Hugging Face | Qwen/Qwen3.8-2.4T-A95B [Hugging Face] |
Qwen3.8 2.4T A95B API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| DeepInfra Qwen/Qwen3.8-2.4T-A95B | $2 | $6 | $0.20 | 262K | models.dev, 4 Oct 2026 |
| SiliconFlow Qwen/Qwen3.8-2.4T-A95B | $2 | $6 | $0.25 | 1M | models.dev, 4 Oct 2026 |
| Fireworks AI accounts/fireworks/models/qwen3p8-2p4t-a95b | $2 | $6 | $0.25 | 262K | models.dev, 4 Oct 2026 |
| Vercel AI Gateway alibaba/qwen3.8-2.4t-a95b | $2 | $6 | $0.25 | 262K | models.dev, 4 Oct 2026 |
| OpenRouter qwen/qwen3.8-2.4t-a95b | $2 | $6 | $0.25 | 1M | OpenRouter, 4 Oct 2026 |
| Hugging Face Inference Providers Qwen/Qwen3.8-2.4T-A95B | $2.50 | $6.25 | – | 262K | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of Qwen3.8 2.4T A95B?
Qwen3.8 2.4T A95B has a context window of 262,144 tokens (262K) and can generate up to 131,072 tokens in a single response, according to models.dev.
How much does the Qwen3.8 2.4T A95B API cost?
Qwen3.8 2.4T A95B is listed from $2 per million input tokens and $6 per million output tokens (DeepInfra). Prices last verified 4 Oct 2026.
Is Qwen3.8 2.4T A95B open source?
Qwen3.8 2.4T A95B is an open-weight model: its weights are published on Hugging Face as Qwen/Qwen3.8-2.4T-A95B under the qwen3.8-max license. Check the license terms before commercial use.
When was Qwen3.8 2.4T A95B released?
Alibaba (Qwen) released Qwen3.8 2.4T A95B on 12 Aug 2026, according to models.dev.
Which providers offer Qwen3.8 2.4T A95B?
We track Qwen3.8 2.4T A95B on 6 providers: DeepInfra, SiliconFlow, Fireworks AI, Vercel AI Gateway, OpenRouter, Hugging Face Inference Providers.
Change history
No changes detected since we started tracking this model. We re-check every source daily.
Sources
- https://openrouter.ai/qwen/qwen3.8-2.4t-a95b
- https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/Qwen/Qwen3.8-2.4T-A95B.toml
- https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B
- https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/Qwen/Qwen3.8-2.4T-A95B.toml
- https://github.com/sst/models.dev/blob/dev/providers/fireworks-ai/models/accounts/fireworks/models/qwen3p8-2p4t-a95b.toml
- https://github.com/sst/models.dev/blob/dev/providers/vercel/models/alibaba/qwen3.8-2.4t-a95b.toml
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/Qwen/Qwen3.8-2.4T-A95B.toml