8/11/2026
What this post added
Introduces DeepSeek-V4-Flash-0731, an open-weight 284B MoE (13B active) model with a 1M context window and selectable reasoning effort, tuned for coding, chat, and agent workflows. Highlights its efficiency-oriented design, hybrid attention architecture for long-context processing, and configurable reasoning-effort levels for trading latency and token usage for deliberation. The 0731 release specifically focuses on agentic performance gains on coding, tool-use, and automation benchmarks.