DEEPSEEK · MODEL RELEASE TRACKER
DeepSeek-V4-Flash
A DeepSeek V4 family model variant (DeepSeek-V4-Flash). Reported as a cost- and compute-efficient V4 variant with 284B total / 13B active parameters, a 1M-token default context, reasoning and agentic performance approaching DeepSeek-V4-Pro, measured benchmark successes on GameASG-Bench and BioPhys-Bridge, and later retired with requests routed to DeepSeek-V4.1-Flash.CURRENT SNAPSHOT4/5 DIMENSIONS WITH DATA
The dimensions that change the decision.
PricingDescribed by DeepSeek as having 'highly cost‑effective API pricing' alongside smaller parameter size and faster response times.
Default context lengthDefault context length is 1,000,000 tokens (1M context is stated as the default across official DeepSeek services).
Not established from the available sources.
BioPhys-Bridge evidence-ID F1In preliminary BioPhys‑Bridge evaluations, DeepSeek‑V4‑Flash obtained the highest evidence‑ID F1 score of 0.360.
GameASG-Bench strict task successesOn GameASG‑Bench, with full tool access and larger nominal turn budgets DeepSeek‑V4‑Flash achieved 18 strict task successes per tested harness; only ten tasks succeeded under both harnesses.
Availability / routingDeepSeek‑V4‑Flash and V4‑Flash‑Vision‑Exp are retired; for compatibility, requests to deepseek‑v4‑flash are temporarily routed to DeepSeek‑V4.1‑Flash.
RELEASE TIMELINE
Published, source-backed release events only.
VERIFIABLE FACTS
Every value stays attached to a source and date.
OTHER · Model size (total / active parameters)DEVELOPER CLAIM
Reported as 284B total parameters and 13B active parameters.
CONTEXT WINDOW · Default context lengthDEVELOPER CLAIM
Default context length is 1,000,000 tokens (1M context is stated as the default across official DeepSeek services).
RELEASE · Release & availabilityDEVELOPER CLAIM
Declared as part of the DeepSeek‑V4 Preview, described as live and open‑sourced, with the API updated and available.
OTHER · Comparative reasoning and agent performanceDEVELOPER CLAIM
Reportedly 'reasoning capabilities closely approach V4‑Pro' and 'performs on par with V4‑Pro on simple Agent tasks' (per DeepSeek).
PRICING · PricingDEVELOPER CLAIM
Described by DeepSeek as having 'highly cost‑effective API pricing' alongside smaller parameter size and faster response times.
OTHER · Retirement and compatibility routingDEVELOPER CLAIM
DeepSeek states 'V4‑Flash & V4‑Flash‑Vision‑Exp are retired'; requests to deepseek‑v4‑flash and deepseek‑v4‑flash‑vision‑exp are temporarily routed to V4.1‑Flash for compatibility.
OTHER · GameASG‑Bench strict task successesINDEPENDENTLY SUPPORTED
On GameASG‑Bench, DeepSeek‑V4‑Flash achieved 18 strict task successes for each tested harness (with full tool access and larger nominal turn budgets), but only 10 tasks succeeded under both harnesses; strict task success rate was reported not monotonic in reasoning effort.
OTHER · BioPhys‑Bridge evidence‑ID F1INDEPENDENTLY SUPPORTED
Reported to obtain the highest evidence‑ID F1 score on BioPhys‑Bridge: 0.360.
CAPABILITIES · Reasoning and agentic performance relative to V4-ProDEVELOPER CLAIM
Reportedly, reasoning capabilities closely approach DeepSeek‑V4‑Pro; performs on par with V4‑Pro on simple agent tasks.
CAPABILITIES · Parameter countsDEVELOPER CLAIM
Reported as 284 billion total parameters and 13 billion active parameters.
BENCHMARKS · BioPhys-Bridge evidence-ID F1INDEPENDENTLY SUPPORTED
In preliminary BioPhys‑Bridge evaluations, DeepSeek‑V4‑Flash obtained the highest evidence‑ID F1 score of 0.360.
BENCHMARKS · GameASG-Bench strict task successesINDEPENDENTLY SUPPORTED
On GameASG‑Bench, with full tool access and larger nominal turn budgets DeepSeek‑V4‑Flash achieved 18 strict task successes per tested harness; only ten tasks succeeded under both harnesses.
AVAILABILITY · Availability / routingDEVELOPER CLAIM
DeepSeek‑V4‑Flash and V4‑Flash‑Vision‑Exp are retired; for compatibility, requests to deepseek‑v4‑flash are temporarily routed to DeepSeek‑V4.1‑Flash.
WHAT CHANGED
Stored passport versions, without reconstructed history.
BENCHMARKS · BioPhys-Bridge evidence-ID F1+ In preliminary BioPhys‑Bridge evaluations, DeepSeek‑V4‑Flash obtained the highest evidence‑ID F1 score of 0.360.
OTHER · BioPhys‑Bridge evidence‑ID F1Reported to obtain the highest evidence‑ID F1 score on BioPhys‑Bridge: 0.360.→Reported to obtain the highest evidence‑ID F1 score on BioPhys‑Bridge: 0.360.
OTHER · Comparative reasoning and agent performanceReportedly 'reasoning capabilities closely approach V4‑Pro' and 'performs on par with V4‑Pro on simple Agent tasks' (per DeepSeek).→Reportedly 'reasoning capabilities closely approach V4‑Pro' and 'performs on par with V4‑Pro on simple Agent tasks' (per DeepSeek).
CONTEXT WINDOW · Default context lengthReported support for a 1,000,000‑token (1M) context as the default across DeepSeek V4 services; model supports dual modes (Thinking / Non‑Thinking).→Default context length is 1,000,000 tokens (1M context is stated as the default across official DeepSeek services).
PRICING · PricingDescribed by DeepSeek as having 'highly cost‑effective API pricing' alongside smaller parameter size and faster response times.→Described by DeepSeek as having 'highly cost‑effective API pricing' alongside smaller parameter size and faster response times.
BENCHMARKS · GameASG-Bench strict task successes+ On GameASG‑Bench, with full tool access and larger nominal turn budgets DeepSeek‑V4‑Flash achieved 18 strict task successes per tested harness; only ten tasks succeeded under both harnesses.
OTHER · GameASG‑Bench strict task successesOn GameASG‑Bench, DeepSeek‑V4‑Flash achieved 18 strict task successes for each tested harness (with full tool access and larger nominal turn budgets), but only 10 tasks succeeded under both harnesses; strict task success rate was reported not monotonic in reasoning effort.→On GameASG‑Bench, DeepSeek‑V4‑Flash achieved 18 strict task successes for each tested harness (with full tool access and larger nominal turn budgets), but only 10 tasks succeeded under both harnesses; strict task success rate was reported not monotonic in reasoning effort.
CAPABILITIES · Parameter counts+ Reported as 284 billion total parameters and 13 billion active parameters.
RELEASE · Release & availabilityDeclared as part of the DeepSeek‑V4 Preview, described as live and open‑sourced, with the API updated and available.→Declared as part of the DeepSeek‑V4 Preview, described as live and open‑sourced, with the API updated and available.
OTHER · Model size (total / active parameters)Reported as 284B total parameters and 13B active parameters.→Reported as 284B total parameters and 13B active parameters.
CAPABILITIES · Reasoning and agentic performance relative to V4-Pro+ Reportedly, reasoning capabilities closely approach DeepSeek‑V4‑Pro; performs on par with V4‑Pro on simple agent tasks.
OTHER · Retirement and compatibility routingDeepSeek states 'V4‑Flash & V4‑Flash‑Vision‑Exp are retired'; requests to deepseek‑v4‑flash and deepseek‑v4‑flash‑vision‑exp are temporarily routed to V4.1‑Flash for compatibility.→DeepSeek states 'V4‑Flash & V4‑Flash‑Vision‑Exp are retired'; requests to deepseek‑v4‑flash and deepseek‑v4‑flash‑vision‑exp are temporarily routed to V4.1‑Flash for compatibility.
AVAILABILITY · Availability / routing+ DeepSeek‑V4‑Flash and V4‑Flash‑Vision‑Exp are retired; for compatibility, requests to deepseek‑v4‑flash are temporarily routed to DeepSeek‑V4.1‑Flash.