DeepSeek-V4.1-Flash
DeepSeek · DeepSeek Flash line · lightweight · open weights
What changed
A new encoder-decoder design that activates only 8B parameters on input and 16B on output, beating V4-Flash on every agentic test DeepSeek published while keeping the MIT licence and 1M-token context — though the API price more than doubles.
Compared with DeepSeek-V4-Flash-0731 · the model it replaces
Finishes 9 in 10 multi-step tasks at the command line
Terminal-Bench 2.1 — running real commands to complete a task
90.6%
▲ up from 82.7% · +8 points
Five points smarter on the all-round score
Artificial Analysis Intelligence Index — a broad average across many tests
40
▲ up from 35 · +5 points
Costs developers more than twice as much to run
Price per 1M tokens — what developers pay, input / output, at peak hours
$0.30 / $1.20
price rose — was $0.14 / $0.28
Where the DeepSeek Flash line stands
DeepSeek-V4.1-Flash is the current model in this line. The card shows how long it has been since the last release against the line's usual gap and its recent ones — the rhythm, not a date.
Every DeepSeek Flash release
Newest first. See the full line with every card →
- Sep 10, 2026DeepSeek-V4.1-Flash · this page
- Jul 31, 2026DeepSeek-V4-Flash-0731
Get the weekly drop
One email a week: what AI models shipped, the actual differences vs the previous version, and what's coming next. No filler.