DeepSeek-V4-Flash
DeepSeek's smaller, fast variant of V4 — same architecture at a fraction of the cost and latency
DeepSeek-V4-Flash
DeepSeek • April 2026
Training Data
Up to early 2026
DeepSeek-V4-Flash
April 2026
Parameters
284 billion (13B active)
Training Method
MoE with hybrid attention, Muon optimizer
Context Window
1,000,000 tokens
Knowledge Cutoff
Not disclosed
Key Features
1M Context • Sparse MoE (13B Active) • Peak/Off-Peak Pricing • Native Responses API • Open Weights (MIT)
Capabilities
Speed: Outstanding
Cost Efficiency: Outstanding
Reasoning: Excellent
What's New in This Version
Left preview on 31 Jul 2026 with agent benchmarks exceeding the V4-Pro preview. Retains V4-Pro's 1M context at roughly a third of the price: $0.22 in / $0.66 out per MTok off-peak, $0.44 / $1.32 at peak
DeepSeek's smaller, fast variant of V4 — same architecture at a fraction of the cost and latency
What's New in This Version
Left preview on 31 Jul 2026 with agent benchmarks exceeding the V4-Pro preview. Retains V4-Pro's 1M context at roughly a third of the price: $0.22 in / $0.66 out per MTok off-peak, $0.44 / $1.32 at peak
Technical Specifications
Key Features
Capabilities
Other DeepSeek Models
Explore more models from DeepSeek
DeepSeek-V4-Pro
DeepSeek's frontier MoE flagship, generally available since August 2026 with substantially stronger agentic tool use and code execution
DeepSeek-V3.2
DeepSeek's latest flagship model matching GPT-5 performance with integrated tool-use thinking
DeepSeek-V3.2-Speciale
DeepSeek's competition-focused variant (EXPIRED Dec 15, 2025 - was temporary API-only release)