Qwen-VL Max
Summary
Qwen-VL Max is a proprietary language model from Alibaba (Qwen), released on 8 Apr 2024. It accepts text and images and generates text, with a 131K-token context window and up to 8K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.80 per million input tokens and $3.20 per million output tokens (Alibaba Cloud Model Studio), across 1 provider we track.
Qwen vision-language model for visual reasoning, documents, and agent tasks
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Alibaba (Qwen) |
|---|---|
| Release date | 8 Apr 2024 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Proprietary (API only) [models.dev] |
| Context window | 131K (131,072 tokens) [models.dev] |
| Max output | 8K tokens [models.dev] |
| Knowledge cutoff | Apr 2024 [models.dev] |
| Input | Text and images [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | No [models.dev] |
| Tool calling | Yes [models.dev] |
Qwen-VL Max API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| Alibaba Cloud Model Studio qwen-vl-max | $0.80 | $3.20 | – | 131K | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of Qwen-VL Max?
Qwen-VL Max has a context window of 131,072 tokens (131K) and can generate up to 8,192 tokens in a single response, according to models.dev.
How much does the Qwen-VL Max API cost?
Through Alibaba Cloud Model Studio, Qwen-VL Max costs $0.80 per million input tokens and $3.20 per million output tokens. Prices last verified 4 Oct 2026.
Is Qwen-VL Max open source?
No. Qwen-VL Max is a proprietary model; Alibaba (Qwen) has not released its weights. It is available through APIs including Alibaba Cloud Model Studio.
When was Qwen-VL Max released?
Alibaba (Qwen) released Qwen-VL Max on 8 Apr 2024, according to models.dev.
Change history
No changes detected since we started tracking this model. We re-check every source daily.