DeepSeek sharply raises prices; V4 Pro launches at 14x its cheapest model

DeepSeek built its reputation on radically cheap inference; its latest move upends that positioning. The flagship V4-Pro-0813 carries 1.7 trillion parameters and launched at roughly 14 times the price of DeepSeek's cheapest model, alongside a new peak/off-peak billing scheme where off-peak requests run at half the peak-hour rate. The price changes took effect August 16.
Technically, V4-Pro is aimed squarely at agentic workloads — tool use, code execution, and long multi-step workflows — and adds compatibility with the OpenAI Responses API format and Codex integration, lowering switching costs for developers already building against OpenAI's interfaces. That agent focus mirrors the industry-wide pivot toward models that can sustain complex, autonomous tasks.
The pricing pivot signals DeepSeek trying to move upmarket into premium territory rather than compete purely on cost — a notable reversal given that its low prices were its defining differentiator. It also opens a gap that rivals are already exploiting: developers on r/DeepSeek pointed out that a $20 ChatGPT Codex subscription running Luna High is now more than 2x cheaper than DeepSeek's old pricing, and others noted alternative providers still serve DeepSeek models more cheaply.
The strategic risk is clear: DeepSeek is asking its most price-sensitive user base to pay premium rates just as edge-deployable open models like Alibaba's Qwen3.8-27B make local, near-free inference more viable. Watch whether V4-Pro's agent performance justifies the premium or whether users migrate to cheaper hosts and open alternatives.