# DeepSeek V4 Flash 0731

> DeepSeek V4 Flash 0731 is an open-weight language model from DeepSeek, released on 31 Jul 2026. It accepts text and generates text, with a 1M-token context window and up to 384K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.08 per million input tokens and $0.15 per million output tokens (Vercel AI Gateway), across 10 providers we track.

Source page: https://www.aimodel.directory/models/deepseek-v4-flash-0731
Last verified: 4 Oct 2026

## Specifications

- **Developer:** DeepSeek
- **Release date:** 31 Jul 2026 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Weights:** Open weights (source: Hugging Face, https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731)
- **License:** mit (source: Hugging Face, https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731)
- **Parameters:** 304B (source: Hugging Face, https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731)
- **Context window:** 1M (1,048,576 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Max output:** 384K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Knowledge cutoff:** May 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Input:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Reasoning mode:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Tool calling:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Structured output:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml)
- **Hugging Face:** deepseek-ai/DeepSeek-V4-Flash-0731 (source: Hugging Face, https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| OpenRouter | $0.02 | $1.28 | $0.02 | 1M |
| DeepInfra | $0.06 | $0.18 | $0.01 | 1M |
| Vercel AI Gateway | $0.08 | $0.15 | $0.01 | 1M |
| Baseten | $0.13 | $0.26 | $0.03 | 1M |
| Nebius AI Studio | $0.14 | $0.28 | $0.14 | 1M |
| Hugging Face Inference Providers | $0.14 | $0.28 | – | 1M |
| Together AI | $0.14 | $0.28 | $0.03 | 1M |
| Alibaba Cloud Model Studio | $0.20 | $0.40 | $0.04 | 1M |
| SiliconFlow | $0.22 | $0.66 | $0.01 | 1M |
| Cloudflare Workers AI | $0.44 | $1.32 | $0.01 | 1M |

## FAQ

### What is the context window of DeepSeek V4 Flash 0731?

DeepSeek V4 Flash 0731 has a context window of 1,048,576 tokens (1M) and can generate up to 384,000 tokens in a single response, according to models.dev.

### How much does the DeepSeek V4 Flash 0731 API cost?

Through Alibaba Cloud Model Studio, DeepSeek V4 Flash 0731 costs $0.20 per million input tokens and $0.40 per million output tokens. The lowest listed price is $0.08 input / $0.15 output through Vercel AI Gateway. Prices last verified 4 Oct 2026.

### Is DeepSeek V4 Flash 0731 open source?

DeepSeek V4 Flash 0731 is an open-weight model: its weights are published on Hugging Face as deepseek-ai/DeepSeek-V4-Flash-0731 under the mit license. Check the license terms before commercial use.

### When was DeepSeek V4 Flash 0731 released?

DeepSeek released DeepSeek V4 Flash 0731 on 31 Jul 2026, according to models.dev.

### Which providers offer DeepSeek V4 Flash 0731?

We track DeepSeek V4 Flash 0731 on 10 providers: OpenRouter, DeepInfra, Vercel AI Gateway, Baseten, Nebius AI Studio, Hugging Face Inference Providers, Together AI, Alibaba Cloud Model Studio, SiliconFlow, Cloudflare Workers AI.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "DeepSeek V4 Flash 0731", https://www.aimodel.directory/models/deepseek-v4-flash-0731