Alibaba (Qwen)

Qwen-VL OCR

Generally availableProprietary34K context

Summary

Qwen-VL OCR is a proprietary language model from Alibaba (Qwen), released on 28 Oct 2024. It accepts text and images and generates text, with a 34K-token context window and up to 4K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.72 per million input tokens and $0.72 per million output tokens (Alibaba Cloud Model Studio), across 1 provider we track.

OCR model for extracting structured text from documents and screenshots

Last verified 4 Oct 2026. Every figure on this page links to its source.

Specifications

DeveloperAlibaba (Qwen)
Release date28 Oct 2024 [models.dev]
StatusGenerally available [models.dev]
WeightsProprietary (API only) [models.dev]
Context window34K (34,096 tokens) [models.dev]
Max output4K tokens [models.dev]
Knowledge cutoffApr 2024 [models.dev]
InputText and images [models.dev]
OutputText [models.dev]
Reasoning modeNo [models.dev]
Tool callingNo [models.dev]

Qwen-VL OCR API pricing by provider

USD per million tokens. Sorted by input price.

ProviderInputOutputCached inputContextSource
Alibaba Cloud Model Studio
qwen-vl-ocr
$0.72$0.72–34Kmodels.dev, 4 Oct 2026

Frequently asked questions

What is the context window of Qwen-VL OCR?

Qwen-VL OCR has a context window of 34,096 tokens (34K) and can generate up to 4,096 tokens in a single response, according to models.dev.

How much does the Qwen-VL OCR API cost?

Through Alibaba Cloud Model Studio, Qwen-VL OCR costs $0.72 per million input tokens and $0.72 per million output tokens. Prices last verified 4 Oct 2026.

Is Qwen-VL OCR open source?

No. Qwen-VL OCR is a proprietary model; Alibaba (Qwen) has not released its weights. It is available through APIs including Alibaba Cloud Model Studio.

When was Qwen-VL OCR released?

Alibaba (Qwen) released Qwen-VL OCR on 28 Oct 2024, according to models.dev.

Change history

No changes detected since we started tracking this model. We re-check every source daily.

Sources