LATEST MODEL
View in the Herd

DeepSeek-V4-Flash

DeepSeek ⚡ The Runner Released April 2026

DeepSeek's smaller, fast variant of V4 — same architecture at a fraction of the cost and latency

DeepSeek-V4-Flash

DeepSeekApril 2026

Latest

Training Data

Up to early 2026

DeepSeek-V4-Flash

April 2026

Parameters

284 billion (13B active)

Training Method

MoE with hybrid attention, Muon optimizer

Context Window

1,000,000 tokens

Knowledge Cutoff

Not disclosed

Key Features

1M Context • Sparse MoE (13B Active) • Peak/Off-Peak Pricing • Native Responses API • Open Weights (MIT)

Capabilities

Speed: Outstanding

Cost Efficiency: Outstanding

Reasoning: Excellent

What's New in This Version

Left preview on 31 Jul 2026 with agent benchmarks exceeding the V4-Pro preview. Retains V4-Pro's 1M context at roughly a third of the price: $0.22 in / $0.66 out per MTok off-peak, $0.44 / $1.32 at peak

DeepSeek's smaller, fast variant of V4 — same architecture at a fraction of the cost and latency

What's New in This Version

Left preview on 31 Jul 2026 with agent benchmarks exceeding the V4-Pro preview. Retains V4-Pro's 1M context at roughly a third of the price: $0.22 in / $0.66 out per MTok off-peak, $0.44 / $1.32 at peak

Technical Specifications

Parameters 284 billion (13B active)
Context Window 1,000,000 tokens
Training Method MoE with hybrid attention, Muon optimizer
Knowledge Cutoff Not disclosed
Training Data Up to early 2026

Key Features

1M Context Sparse MoE (13B Active) Peak/Off-Peak Pricing Native Responses API Open Weights (MIT)

Capabilities

Speed: Outstanding
Cost Efficiency: Outstanding
Reasoning: Excellent
Theme
Language
Support
© funclosure 2025