Tech Meridian ← ENTITY INDEX
RU

MODEL · ENTITY #3475

DeepSeek Sparse Attention (DSA)

Related event timeline, sources and context from the news index.

EVENT TIMELINE

2

MODELS · 1 SOURCE · DeepSeek

DeepSeek publishes open-source DeepSeek‑V4 preview with 1M-token context

DeepSeek has open-sourced a preview of DeepSeek‑V4, offering two variants: DeepSeek‑V4‑Pro (1.6T total / 49B active params) and DeepSeek‑V4‑Flash (284B total / 13B active params). The company says both models support a default 1M-token context, introduce novel token-wise compression and DSA (DeepSeek Sparse Attention), claim open-source state‑of‑the‑art agentic and reasoning performance (trailing only Gemini‑3.1‑Pro among models they compare to), and are available now via chat.deepseek.com and updated APIs with compatibility for OpenAI ChatCompletions & Anthropic endpoints; older models deepseek-chat and deepseek-reasoner will be retired on Jul 24, 2026 at 15:59 UTC.

8.0

MODELS · 1 SOURCE · DeepSeek

DeepSeek launches experimental model V3.2-Exp with sparse attention and big API price cuts

DeepSeek introduced DeepSeek-V3.2-Exp, an experimental model built on V3.1-Terminus that debuts DeepSeek Sparse Attention (DSA) to improve long-context efficiency with minimal quality loss; published benchmarks show parity with V3.1-Terminus. DeepSeek also cut DeepSeek API prices by over 50% effective immediately, provides TileLang and CUDA GPU kernels, and keeps V3.1-Terminus available via a temporary API until Oct 15, 2025, 15:59 UTC for comparison testing.

7.0