Back
AlibabaAugust 03, 20262 sources

Alibaba unveils Qwen3.8-Max at 2.4 trillion parameters, claims parity with Western frontier models

AI Analysis

Alibaba's Qwen team launched Qwen3.8-Max, calling it their 'most capable' model, with 2.4 trillion total parameters and up to 95 billion active per request across a multimodal MoE architecture supporting text, image, and video input and a 1 million token context window. In benchmarking Alibaba claims it beats Moonshot's Kimi K3 and rivals OpenAI and Anthropic's frontier offerings, and it reportedly autonomously executed a 16-day software-engineering project during testing. Pricing lands at roughly $2/M input and $6/M output tokens — about one-fifth the cost of comparable Western models — and it is generally available via QwenCloud and Alibaba Cloud Model Studio, alongside a new QwenWork enterprise agent product.

Crucially, Alibaba plans an open-weights release of Qwen3.8-Max and a laptop-sized Qwen3.8-27B on August 10. The 27B variant already drew huge developer excitement: Unsloth's Daniel Han validated it will run in just 17GB of VRAM, and Simon Willison said he's 'very much looking forward' to the laptop-sized models. The r/LocalLLaMA thread announcing the pair drew nearly 2,700 upvotes.

The release feeds this week's dominant theme of Chinese labs pressing an aggressive cost-and-openness advantage. Shares rallied on the news. Skeptics caution that vendor benchmarks arrive fast and independent verification is still pending — a recurring critique of first-party claims of frontier parity. Watch the August 10 weights drop and third-party evals to see whether the parity claim survives scrutiny.

Sources
AI Briefing
·Vendors·Curated by AI agents · Updated daily · 2026
Built by Koby Almog