DeepSeek V4.1 Flash
High-volume coding and text workloads where budget is the top priority.
DeepSeek V4.1 Flash supersedes DeepSeek V4 Flash.
What are DeepSeek V4.1 Flash's specs and price?
DeepSeek V4.1 Flash, built by DeepSeek, ships a 1M-token context window and a 384K-token max output, released 2026-09. It supports text and vision input with a dedicated reasoning mode and costs $0.52 per million blended tokens, the 14th-cheapest of 49 models we track.
What are DeepSeek V4.1 Flash's specs?
| Context window | 1M tokens |
| Max output | 384K tokens |
| Modalities | text, vision |
| Extended thinking | Yes |
| Released | 2026-09 |
| Knowledge cutoff | Not published |
| Provider | DeepSeek |
Verified 2026-10-01 — source.
Where does DeepSeek V4.1 Flash rank?
What are DeepSeek V4.1 Flash's strengths?
- Open-weights MoE flagship with native vision
- Thinking mode on by default
- Off-peak pricing undercuts every Western budget model
What else should you know about DeepSeek V4.1 Flash?
What are common questions about DeepSeek V4.1 Flash?
What is DeepSeek V4.1 Flash's context window?
DeepSeek V4.1 Flash has a 1M-token context window and a 384K-token max output — the 19th-largest context of the 49 current models we track. Source: https://api-docs.deepseek.com/quick_start/pricing, verified 2026-10-01.
Does DeepSeek V4.1 Flash support vision or audio input?
Yes — DeepSeek V4.1 Flash accepts vision input in addition to text.
Does DeepSeek V4.1 Flash have a reasoning or extended-thinking mode?
Yes — DeepSeek V4.1 Flash exposes a dedicated reasoning mode for multi-step problems.
When was DeepSeek V4.1 Flash released, and what is its knowledge cutoff?
DeepSeek V4.1 Flash was released 2026-09.
How much does DeepSeek V4.1 Flash cost, and who provides it?
DeepSeek V4.1 Flash is served by DeepSeek at $0.52/M blended tokens (3:1 input:output) — the 14th-cheapest of 49 current models. Full pricing breakdown: /llm-api-pricing/deepseek-v4-1-flash.
