Alibaba's Qwen Audio 3.0 tops OpenAI on speech; Qwen3.8-Max previewed

Alibaba pushed on two fronts at the World AI Conference in Shanghai. Qwen-Audio-3.0-Realtime enables live two-way voice conversations and, per Alibaba, beat OpenAI on a new speech benchmark on July 28 — a direct challenge in the real-time voice segment where xAI also just claimed wins. Qwen already powers an ecosystem-integrated personal assistant across Taobao and Alipay, giving Alibaba a massive consumer distribution surface for the audio model.
On the frontier side, Alibaba previewed Qwen3.8-Max, a 2.4-trillion-parameter multimodal model it says is 'second only to Anthropic's Fable 5' — though notably without releasing supporting benchmark data, and with open weights still pending. Preview access runs through Alibaba's Token Plan, Qoder, and QoderWork platforms, with an open-weight release promised 'soon.' Alibaba is launching a #QwenGrowthPlan to get developers testing Qwen3.8's agentic capabilities on real-world tasks.
Community skepticism is pointed: developers note 'preview' models with unreleased weights aren't truly open or free to run, and an r/LocalLLaMA thread ('Is there really nothing better under 120B?') showed enduring loyalty to existing Qwen models. Still, Chinese labs — Qwen, DeepSeek V4, GLM-5.2 — are collectively pressuring US incumbents on price and openness. What to watch: whether Qwen3.8-Max ships open weights with real benchmarks, and how Qwen Audio 3.0 fares against OpenAI and xAI voice models in independent testing.