DeepSeek V4.1 Flash
Summary
DeepSeek V4.1 Flash is an open-weight language model from DeepSeek, released on 10 Sept 2026. It accepts text and images and generates text, with a 1M-token context window and up to 393K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.15 per million input tokens and $0.60 per million output tokens (DeepSeek API), across 10 providers we track.
DeepSeek V4.1 Flash model for reasoning and agentic coding
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | DeepSeek |
|---|---|
| Release date | 10 Sept 2026 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | mit [Hugging Face] |
| Parameters | 763B [Hugging Face] |
| Context window | 1M (1,000,000 tokens) [models.dev] |
| Max output | 393K tokens [models.dev] |
| Knowledge cutoff | May 2025 [models.dev] |
| Input | Text and images [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Structured output | Yes [models.dev] |
| Hugging Face | deepseek-ai/DeepSeek-V4.1-Flash [Hugging Face] |
DeepSeek V4.1 Flash API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| NVIDIA NIM deepseek-ai/deepseek-v4.1-flash | Free | Free | – | 1M | models.dev, 4 Oct 2026 |
| OpenRouter deepseek/deepseek-v4.1-flash | $0.003 | $2.40 | $0.003 | 1M | OpenRouter, 4 Oct 2026 |
| DeepSeek API deepseek-flash | $0.15 | $0.60 | $0.003 | 1M | models.dev, 4 Oct 2026 |
| DeepInfra deepseek-ai/DeepSeek-V4.1-Flash | $0.20 | $0.60 | $0.006 | 1M | models.dev, 4 Oct 2026 |
| Baseten deepseek-ai/DeepSeek-V4.1-Flash | $0.30 | $1.20 | $0.03 | 1M | models.dev, 4 Oct 2026 |
| Nebius AI Studio deepseek-ai/DeepSeek-V4.1-Flash | $0.30 | $1.20 | $0.30 | 1M | models.dev, 4 Oct 2026 |
| Fireworks AI accounts/fireworks/routers/deepseek-flash-latest | $0.30 | $1.20 | $0.006 | 1M | models.dev, 4 Oct 2026 |
| Together AI deepseek-ai/DeepSeek-V4.1-Flash | $0.30 | $1.20 | $0.006 | 1M | models.dev, 4 Oct 2026 |
| Hugging Face Inference Providers deepseek-ai/DeepSeek-V4.1-Flash | $0.30 | $1.20 | – | 1M | models.dev, 4 Oct 2026 |
| Fireworks AI accounts/fireworks/models/deepseek-v4p1-flash | $0.30 | $1.20 | $0.006 | 1M | models.dev, 4 Oct 2026 |
| Vercel AI Gateway deepseek/deepseek-v4.1-flash | $0.30 | $1.20 | $0.007 | 1M | models.dev, 4 Oct 2026 |
| Vercel AI Gateway deepseek/deepseek-v4.1-flash-fast | $0.30 | $1.20 | $0.006 | 1M | models.dev, 4 Oct 2026 |
| Baseten deepseek-ai/DeepSeek-V4.1-Flash-Fast | $0.60 | $2.40 | – | 1M | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of DeepSeek V4.1 Flash?
DeepSeek V4.1 Flash has a context window of 1,000,000 tokens (1M) and can generate up to 393,216 tokens in a single response, according to models.dev.
How much does the DeepSeek V4.1 Flash API cost?
Through DeepSeek API, DeepSeek V4.1 Flash costs $0.15 per million input tokens and $0.60 per million output tokens. Prices last verified 4 Oct 2026.
Is DeepSeek V4.1 Flash open source?
DeepSeek V4.1 Flash is an open-weight model: its weights are published on Hugging Face as deepseek-ai/DeepSeek-V4.1-Flash under the mit license. Check the license terms before commercial use.
When was DeepSeek V4.1 Flash released?
DeepSeek released DeepSeek V4.1 Flash on 10 Sept 2026, according to models.dev.
Which providers offer DeepSeek V4.1 Flash?
We track DeepSeek V4.1 Flash on 10 providers: NVIDIA NIM, OpenRouter, DeepSeek API, DeepInfra, Baseten, Nebius AI Studio, Fireworks AI, Together AI, Hugging Face Inference Providers, Vercel AI Gateway.
Change history
No changes detected since we started tracking this model. We re-check every source daily.
Sources
- https://openrouter.ai/deepseek/deepseek-v4.1-flash
- https://github.com/sst/models.dev/blob/dev/providers/deepseek/models/deepseek-flash.toml
- https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash
- https://github.com/sst/models.dev/blob/dev/providers/nvidia/models/deepseek-ai/deepseek-v4.1-flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/deepseek-ai/DeepSeek-V4.1-Flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/baseten/models/deepseek-ai/DeepSeek-V4.1-Flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/nebius/models/deepseek-ai/DeepSeek-V4.1-Flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/fireworks-ai/models/accounts/fireworks/routers/deepseek-flash-latest.toml
- https://github.com/sst/models.dev/blob/dev/providers/togetherai/models/deepseek-ai/DeepSeek-V4.1-Flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/deepseek-ai/DeepSeek-V4.1-Flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/fireworks-ai/models/accounts/fireworks/models/deepseek-v4p1-flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/vercel/models/deepseek/deepseek-v4.1-flash.toml
- https://github.com/sst/models.dev/blob/dev/providers/vercel/models/deepseek/deepseek-v4.1-flash-fast.toml
- https://github.com/sst/models.dev/blob/dev/providers/baseten/models/deepseek-ai/DeepSeek-V4.1-Flash-Fast.toml