All changelog entries
FeatureNew0.14.0ModelsVision

DeepSeek-V4.1-Flash — a big step up for agents and terminals

September 18, 2026

DeepSeek-V4.1-Flash is live. It reads text and images, carries a 1M-token context window, and is built for input-heavy agentic and coding work. Reasoning effort is adjustable, so it can answer quickly or think hard on request. Call it as DeepSeek-V4.1-Flash.

DeepSeek publishes these scores against the earlier V4 Flash, all at maximum reasoning effort:

BenchmarkV4.1 FlashV4 Flash
Terminal-Bench 2.190.682.7
Terminal-Bench 3.030.07.6
DeepSWE v1.1 (resolved)74.254.4
NL2Repo-Bench64.054.2
SEC-Bench Pro62.830.9
HLE with tools63.951.5
Codeforces rating34713289
GPQA Diamond90.989.9

It costs 2 requests of quota per call and is available on every paid plan and the free tier. DeepSeek-V4-Flash stays exactly as it is at 1 request per call, so nothing changes for existing setups. Pick V4.1 when the task is a long agent run or terminal work, and V4 when you want the cheapest fast model.

Source: DeepSeek-V4.1-Flash model card.