DeepSeek V4 Flash 0731
Summary
DeepSeek V4 Flash 0731 is an open-weight language model from DeepSeek, released on 31 Jul 2026. It accepts text and generates text, with a 1M-token context window and up to 384K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.08 per million input tokens and $0.15 per million output tokens (Vercel AI Gateway), across 10 providers we track.
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | DeepSeek |
|---|---|
| Release date | 31 Jul 2026 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | mit [Hugging Face] |
| Parameters | 304B [Hugging Face] |
| Context window | 1M (1,048,576 tokens) [models.dev] |
| Max output | 384K tokens [models.dev] |
| Knowledge cutoff | May 2025 [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Structured output | Yes [models.dev] |
| Hugging Face | deepseek-ai/DeepSeek-V4-Flash-0731 [Hugging Face] |
DeepSeek V4 Flash 0731 API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| OpenRouter deepseek/deepseek-v4-flash-0731 | $0.02 | $1.28 | $0.02 | 1M | OpenRouter, 4 Oct 2026 |
| DeepInfra deepseek-ai/DeepSeek-V4-Flash-0731 | $0.06 | $0.18 | $0.01 | 1M | models.dev, 4 Oct 2026 |
| Vercel AI Gateway deepseek/deepseek-v4-flash-0731 | $0.08 | $0.15 | $0.01 | 1M | models.dev, 4 Oct 2026 |
| Baseten deepseek-ai/DeepSeek-V4-Flash-0731 | $0.13 | $0.26 | $0.03 | 1M | models.dev, 4 Oct 2026 |
| Nebius AI Studio deepseek-ai/DeepSeek-V4-Flash-0731 | $0.14 | $0.28 | $0.14 | 1M | models.dev, 4 Oct 2026 |
| Hugging Face Inference Providers deepseek-ai/DeepSeek-V4-Flash-0731 | $0.14 | $0.28 | – | 1M | models.dev, 4 Oct 2026 |
| Together AI deepseek-ai/DeepSeek-V4-Flash-0731 | $0.14 | $0.28 | $0.03 | 1M | models.dev, 4 Oct 2026 |
| Alibaba Cloud Model Studio deepseek-v4-flash-0731 | $0.20 | $0.40 | $0.04 | 1M | models.dev, 4 Oct 2026 |
| SiliconFlow deepseek-ai/DeepSeek-V4-Flash-0731 | $0.22 | $0.66 | $0.01 | 1M | models.dev, 4 Oct 2026 |
| Cloudflare Workers AI @cf/deepseek-ai/deepseek-v4-flash-0731 | $0.44 | $1.32 | $0.01 | 1M | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of DeepSeek V4 Flash 0731?
DeepSeek V4 Flash 0731 has a context window of 1,048,576 tokens (1M) and can generate up to 384,000 tokens in a single response, according to models.dev.
How much does the DeepSeek V4 Flash 0731 API cost?
Through Alibaba Cloud Model Studio, DeepSeek V4 Flash 0731 costs $0.20 per million input tokens and $0.40 per million output tokens. The lowest listed price is $0.08 input / $0.15 output through Vercel AI Gateway. Prices last verified 4 Oct 2026.
Is DeepSeek V4 Flash 0731 open source?
DeepSeek V4 Flash 0731 is an open-weight model: its weights are published on Hugging Face as deepseek-ai/DeepSeek-V4-Flash-0731 under the mit license. Check the license terms before commercial use.
When was DeepSeek V4 Flash 0731 released?
DeepSeek released DeepSeek V4 Flash 0731 on 31 Jul 2026, according to models.dev.
Which providers offer DeepSeek V4 Flash 0731?
We track DeepSeek V4 Flash 0731 on 10 providers: OpenRouter, DeepInfra, Vercel AI Gateway, Baseten, Nebius AI Studio, Hugging Face Inference Providers, Together AI, Alibaba Cloud Model Studio, SiliconFlow, Cloudflare Workers AI.
Change history
No changes detected since we started tracking this model. We re-check every source daily.
Sources
- https://openrouter.ai/deepseek/deepseek-v4-flash-0731
- https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml
- https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
- https://github.com/sst/models.dev/blob/dev/providers/vercel/models/deepseek/deepseek-v4-flash-0731.toml
- https://github.com/sst/models.dev/blob/dev/providers/baseten/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml
- https://github.com/sst/models.dev/blob/dev/providers/nebius/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml
- https://github.com/sst/models.dev/blob/dev/providers/togetherai/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml
- https://github.com/sst/models.dev/blob/dev/providers/alibaba/models/deepseek-v4-flash-0731.toml
- https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/deepseek-ai/DeepSeek-V4-Flash-0731.toml
- https://github.com/sst/models.dev/blob/dev/providers/cloudflare-workers-ai/models/@cf/deepseek-ai/deepseek-v4-flash-0731.toml