This page may contain stale information. Last updated: 2026-08-14

Overview

DeepSeek’s efficiency-oriented MoE model in the V4 family: 284B total / 13B active parameters, 1M context, MIT weights. Official public-beta graduation as DeepSeek-V4-Flash-0731 on July 31, 2026 (API id deepseek-v4-flash).

Recent Developments

  • 2026-08-16: Same peak/off-peak pricing model applies; illustrative costs rise ~2× off-peak / ~4× peak vs flat (2026-08-13-deepseek-api-pricing-docs)

  • 2026-07-31: Preview → official; architecture unchanged; improvements via re-post-training + dspark speculative decoding; native Responses API; Codex adaptation (2026-08-01-deepseek-v4-flash-0731-seawork)

  • Vendor claims agent benchmarks exceed V4-Pro Preview at ⅓ Pro output price (0.28 per M)

Warning

Benchmark wins are vendor-reported on DeepSeek’s unreleased harness — not independently verified.

Sources