Alibaba Unveils 2.4-Trillion-Parameter Qwen3.8-Max, Claims Parity With Anthropic Fable 5

Alibaba made Qwen3.8-Max generally available on August 3 via QwenCloud and Alibaba Cloud Model Studio, positioning it as its new flagship. The Mixture-of-Experts model spans 2.4 trillion total parameters, activating roughly 95 billion per query, and supports text, image, and video input with a 1 million token context window. Internal tests claim 125+ hours of sustained autonomous execution alongside leading visual-reasoning and agentic computer-use benchmarks.
Pricing is the headline for enterprises: $2.00 per million input tokens and $6.00 per million output tokens, a fraction of Anthropic's $10–$50 range for comparable frontier capability. Open weights are promised for the week of August 10, which — combined with the price — extends the same China-price-war pressure DeepSeek is applying, but at the high end of capability rather than the low end.
The market noticed: Alibaba's Hong Kong-listed shares rallied about 6% on the news. Alibaba also consolidated three workplace-agent products into a new enterprise offering, QwenWork, signaling a push to monetize agents in the office productivity layer where Microsoft Copilot and Google Workspace compete.
Competitive context matters here. Independent AI commentator Simon Willison said he was 'very much looking forward to the upcoming laptop-sized Qwen 3.8 models,' underscoring that the smaller open-weight variants may matter as much as the flagship for the local-inference crowd. The skeptical read, echoed across developer forums this week: benchmark parity claims from a vendor's own internal tests deserve scrutiny until third parties and the open weights arrive. But the strategic thrust is unambiguous — Chinese labs are now contesting the frontier on capability, not just undercutting on price, and doing it with open weights that Western closed labs won't match.