# Qwen3 VL 8B Instruct

> Qwen3 VL 8B Instruct is an open-weight language model from Alibaba (Qwen), released on 14 Oct 2025. It accepts images and text and generates text, with a 262K-token context window and up to 33K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.12 per million input tokens and $0.46 per million output tokens (OpenRouter), across 1 provider we track. It is scheduled for retirement on 9 Oct 2026.

Source page: https://www.aimodel.directory/models/qwen3-vl-8b-instruct
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Alibaba (Qwen)
- **Release date:** 14 Oct 2025 (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Status:** Generally available
- **Weights:** Open weights (source: Hugging Face, https://huggingface.co/Qwen/Qwen3-VL-8B-Instruct)
- **License:** apache-2.0 (source: Hugging Face, https://huggingface.co/Qwen/Qwen3-VL-8B-Instruct)
- **Parameters:** 8.8B (source: Hugging Face, https://huggingface.co/Qwen/Qwen3-VL-8B-Instruct)
- **Context window:** 262K (262,144 tokens) (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Max output:** 33K tokens (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Input:** Images and text (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Output:** Text (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Reasoning mode:** No (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Tool calling:** Yes (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Structured output:** Yes (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Retirement date:** 9 Oct 2026 (source: OpenRouter, https://openrouter.ai/qwen/qwen3-vl-8b-instruct)
- **Hugging Face:** Qwen/Qwen3-VL-8B-Instruct (source: Hugging Face, https://huggingface.co/Qwen/Qwen3-VL-8B-Instruct)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| OpenRouter | $0.12 | $0.46 | – | 131K |

## FAQ

### What is the context window of Qwen3 VL 8B Instruct?

Qwen3 VL 8B Instruct has a context window of 262,144 tokens (262K) and can generate up to 32,768 tokens in a single response, according to OpenRouter.

### How much does the Qwen3 VL 8B Instruct API cost?

Qwen3 VL 8B Instruct is listed from $0.12 per million input tokens and $0.46 per million output tokens (OpenRouter). Prices last verified 4 Oct 2026.

### Is Qwen3 VL 8B Instruct open source?

Qwen3 VL 8B Instruct is an open-weight model: its weights are published on Hugging Face as Qwen/Qwen3-VL-8B-Instruct under the apache-2.0 license. Check the license terms before commercial use.

### When was Qwen3 VL 8B Instruct released?

Alibaba (Qwen) released Qwen3 VL 8B Instruct on 14 Oct 2025, according to OpenRouter.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Qwen3 VL 8B Instruct", https://www.aimodel.directory/models/qwen3-vl-8b-instruct