DeepSeek V4 goes GA, retires legacy aliases and adds surge pricing

DeepSeek moved its V4 model family to general availability and cleaned up its API lineup: the legacy `deepseek-chat` and `deepseek-reasoner` aliases were retired on July 24 and now redirect to `deepseek-v4-flash` or `deepseek-v4-pro`, forcing a migration for existing integrations. The company also confirmed it will introduce time-of-day surge pricing — an unusual move that ties inference cost to demand.
The V4 split is capability-tiered: V4 Flash is optimized for speed and cost-efficiency on chat and classification workloads, while V4 Pro handles complex reasoning and long-context tasks up to 1 million tokens. This positions DeepSeek to compete on both the cheap-and-fast and heavy-reasoning ends simultaneously.
The launch fits the week's dominant theme — Chinese open models pressuring US frontier pricing — and comes as DeepSeek's founder Liang Wenfeng said the company prioritizes long-term AGI development over profit, plans to keep its most advanced models open-source, is developing its own inference chip to cut NVIDIA dependence, and is weighing a 2026 STAR Market IPO at a $71–74B valuation. A viral r/DeepSeek thread (452 upvotes) noted the CEO framing '6x profit' as corporate 'restraint,' surprising users who assumed the company was burning cash for share. Skeptics should watch how surge pricing lands with cost-sensitive developers and whether the alias retirement breaks production integrations. The open-source commitment also makes DeepSeek a live example in the open-weight policy debate.