Back
DeepSeekAugust 04, 20261 sources

DeepSeek V4-Flash launches at 28 cents per million output tokens, tops global usage

AI Analysis

DeepSeek officially released V4-Flash, an ultra-low-cost model with enhanced agent capabilities priced at $0.14 per million input tokens and $0.28 per million output tokens. Artificial Analysis / a research firm cited by Reuters estimated V4-Flash averages about 3 cents per benchmark test, making it by far the cheapest well-known global model to run — reportedly more than 100 times cheaper than Anthropic's Claude Fable 5. DeepSeek also made a 75% price cut on V4-Pro permanent.

Within a day, V4-Flash had topped global token usage charts at roughly 7.1 trillion tokens weekly, and CGTN reported nine of the top 10 most-used models are now Chinese. Its benchmark performance trails the most powerful Western alternatives, reinforcing the framing that Chinese labs are competing on cost-efficiency rather than raw capability — a debate that dominated developer forums this week.

Finance and AI developers hailed the pricing as a 'new economic baseline' one to two orders of magnitude below Western flagships. On r/DeepSeek, users showcased a full medical website built with V4-Flash (627 upvotes) and debated the cheapest access routes. The launch, arriving the same week as Alibaba's Qwen3.8-Max, cements a China price war that is resetting the entire industry's cost expectations. The open question: whether Western labs respond with matching cuts or lean harder on capability differentiation.

Sources
AI Briefing
·Vendors·Curated by AI agents · Updated daily · 2026
Built by Koby Almog