DeepSeek V4-Flash Adds Vision at No Price Premium
DeepSeek added image understanding to its cheapest fast model on August 21, 2026, billing images at standard V4-Flash rates with a free Files API.
DeepSeek added image understanding to its cheapest fast model on August 21, 2026, billing images at standard V4-Flash rates with a free Files API.
DeepSeek pushed its flagship V4 Pro to general availability on August 12, 2026, ending a preview that ran nearly four months. The GA build is a 1.6 trillion parameter MoE with a 1M-token context window.
DeepSeek shipped the official production build of V4 Flash on July 31, 2026: an open-weight, MIT-licensed 284B-parameter MoE model with a 1M token context, native OpenAI Responses API and Codex support, and API output at $0.28 per million tokens.
An open-source DeepSeek-native coding agent built around prefix-cache stability. A 435M-token session that costs $61 on Claude runs $12 on Reasonix.
DeepSeek locked V4-Pro at $0.435 input miss, $0.003625 cache hit, $0.87 output per million tokens on May 23, ending the May 31 expiry and resetting the price floor for frontier reasoning APIs.
DeepSeek-V4-Flash is the first local model competitive with frontier AI, making LLM activation steering practical for the first time. A guide for creators.