DeepSeek-V4-Flash-0731

DeepSeek's MIT-licensed 284B MoE model, retrained for stronger agentic coding

Free Reasoning
Visit Product Page →
1000000 context tokens
284B (13B active) parameters
MIT license
Jul 2026 released

DeepSeek-V4-Flash-0731 moves DeepSeek-V4-Flash out of preview and into public beta as a re-post-trained checkpoint, announced July 31, 2026 in the company’s API changelog. The architecture is unchanged from the model’s April debut, a 284 billion parameter mixture-of-experts design with 13 billion active parameters per token, a shared expert plus 256 routed experts, and a 1 million token context window. What changed is the post-training pipeline: DeepSeek retuned it specifically for coding, agentic tool use, and reasoning, and the API now natively supports the Responses format and works with Codex out of the box. It remains MIT-licensed and ungated, so teams can self-host it commercially without restriction, with roughly 110 GB of memory needed at 3-bit quantization or a 4x GB300 node for full precision. API pricing stayed low at $0.14 per million input tokens and $0.28 per million output tokens.