Tech Meridian← ALL MODELS
PROMY MERIDIAN

DEEPSEEK · MODEL RELEASE TRACKER

DeepSeek-V4-Flash

A DeepSeek V4 family model variant (DeepSeek-V4-Flash). Reported as a cost- and compute-efficient V4 variant with 284B total / 13B active parameters, a 1M-token default context, reasoning and agentic performance approaching DeepSeek-V4-Pro, measured benchmark successes on GameASG-Bench and BioPhys-Bridge, and later retired with requests routed to DeepSeek-V4.1-Flash.

CURRENT SNAPSHOT4/5 DIMENSIONS WITH DATA

The dimensions that change the decision.

PRICING
PricingDescribed by DeepSeek as having 'highly cost‑effective API pricing' alongside smaller parameter size and faster response times.
CONTEXT WINDOW
Default context lengthDefault context length is 1,000,000 tokens (1M context is stated as the default across official DeepSeek services).
MODALITIES

Not established from the available sources.

BENCHMARKS
BioPhys-Bridge evidence-ID F1In preliminary BioPhys‑Bridge evaluations, DeepSeek‑V4‑Flash obtained the highest evidence‑ID F1 score of 0.360.
GameASG-Bench strict task successesOn GameASG‑Bench, with full tool access and larger nominal turn budgets DeepSeek‑V4‑Flash achieved 18 strict task successes per tested harness; only ten tasks succeeded under both harnesses.
AVAILABILITY
Availability / routingDeepSeek‑V4‑Flash and V4‑Flash‑Vision‑Exp are retired; for compatibility, requests to deepseek‑v4‑flash are temporarily routed to DeepSeek‑V4.1‑Flash.

RELEASE TIMELINE

Published, source-backed release events only.

1 SOURCES · IMPORTANCE 8.0

DeepSeek launches DeepSeek‑V4.1‑Flash with new causal encoder–decoder and lower costs

DeepSeek announced DeepSeek‑V4.1‑Flash, a smaller-native-vision model using a new Causal Encoder–Decoder design (8B active params for input, 16B for output), new pretraining and larger-scale RL post-training, plus KV‑cache compression. The company is retiring V4‑Flash variants, routing deepseek-v4-pro traffic to V4.1‑Flash from 04:00 UTC on Sept 14, 2026, and changing pricing (new rates take effect 04:00 UTC on Sept 10, 2026) while partners WorkBuddy, CodeBuddy and OpenCode already support V4.1‑Flash.

→

VERIFIABLE FACTS

Every value stays attached to a source and date.

WHAT CHANGED

Stored passport versions, without reconstructed history.

Passport updated13 facts
BENCHMARKS · BioPhys-Bridge evidence-ID F1+ In preliminary BioPhys‑Bridge evaluations, DeepSeek‑V4‑Flash obtained the highest evidence‑ID F1 score of 0.360.
OTHER · BioPhys‑Bridge evidence‑ID F1Reported to obtain the highest evidence‑ID F1 score on BioPhys‑Bridge: 0.360.→Reported to obtain the highest evidence‑ID F1 score on BioPhys‑Bridge: 0.360.
OTHER · Comparative reasoning and agent performanceReportedly 'reasoning capabilities closely approach V4‑Pro' and 'performs on par with V4‑Pro on simple Agent tasks' (per DeepSeek).→Reportedly 'reasoning capabilities closely approach V4‑Pro' and 'performs on par with V4‑Pro on simple Agent tasks' (per DeepSeek).
CONTEXT WINDOW · Default context lengthReported support for a 1,000,000‑token (1M) context as the default across DeepSeek V4 services; model supports dual modes (Thinking / Non‑Thinking).→Default context length is 1,000,000 tokens (1M context is stated as the default across official DeepSeek services).
PRICING · PricingDescribed by DeepSeek as having 'highly cost‑effective API pricing' alongside smaller parameter size and faster response times.→Described by DeepSeek as having 'highly cost‑effective API pricing' alongside smaller parameter size and faster response times.
BENCHMARKS · GameASG-Bench strict task successes+ On GameASG‑Bench, with full tool access and larger nominal turn budgets DeepSeek‑V4‑Flash achieved 18 strict task successes per tested harness; only ten tasks succeeded under both harnesses.
OTHER · GameASG‑Bench strict task successesOn GameASG‑Bench, DeepSeek‑V4‑Flash achieved 18 strict task successes for each tested harness (with full tool access and larger nominal turn budgets), but only 10 tasks succeeded under both harnesses; strict task success rate was reported not monotonic in reasoning effort.→On GameASG‑Bench, DeepSeek‑V4‑Flash achieved 18 strict task successes for each tested harness (with full tool access and larger nominal turn budgets), but only 10 tasks succeeded under both harnesses; strict task success rate was reported not monotonic in reasoning effort.
CAPABILITIES · Parameter counts+ Reported as 284 billion total parameters and 13 billion active parameters.
RELEASE · Release & availabilityDeclared as part of the DeepSeek‑V4 Preview, described as live and open‑sourced, with the API updated and available.→Declared as part of the DeepSeek‑V4 Preview, described as live and open‑sourced, with the API updated and available.
OTHER · Model size (total / active parameters)Reported as 284B total parameters and 13B active parameters.→Reported as 284B total parameters and 13B active parameters.
CAPABILITIES · Reasoning and agentic performance relative to V4-Pro+ Reportedly, reasoning capabilities closely approach DeepSeek‑V4‑Pro; performs on par with V4‑Pro on simple agent tasks.
OTHER · Retirement and compatibility routingDeepSeek states 'V4‑Flash & V4‑Flash‑Vision‑Exp are retired'; requests to deepseek‑v4‑flash and deepseek‑v4‑flash‑vision‑exp are temporarily routed to V4.1‑Flash for compatibility.→DeepSeek states 'V4‑Flash & V4‑Flash‑Vision‑Exp are retired'; requests to deepseek‑v4‑flash and deepseek‑v4‑flash‑vision‑exp are temporarily routed to V4.1‑Flash for compatibility.
AVAILABILITY · Availability / routing+ DeepSeek‑V4‑Flash and V4‑Flash‑Vision‑Exp are retired; for compatibility, requests to deepseek‑v4‑flash are temporarily routed to DeepSeek‑V4.1‑Flash.
Passport created8 facts