← All vendors
Alibaba logo
Vendor

Alibaba AI News

Every AI news story AI Briefing has published about Alibaba — 86 articles spanning Apr 13, 2026 – Aug 5, 2026. Track Alibaba's model releases, research papers, product launches, funding rounds, and partnerships across the AI industry, updated daily.

86 articles · Apr 13, 2026 – Aug 5, 2026

Alibaba unveils Qwen3.8-Max at 2.4 trillion parameters, claims parity with Western frontier models

Alibaba released Qwen3.8-Max, its most capable model yet, a 2.4-trillion-parameter Mixture-of-Experts model activating up to 95B parameters per query, priced at $2/M input and $6/M output tokens — about one-fifth the cost of comparable Western models. It supports a 1M-token context window, and open weights for Qwen3.8-Max plus a smaller Qwen3.8-27B are due August 10.

2026-08-05

Alibaba Launches Qwen 3.8-Max at 2.4 Trillion Parameters, Stock Jumps 7%

Alibaba unveiled Qwen 3.8-Max, its largest and most capable model to date at 2.4 trillion parameters with a 1-million-token context window, natively multimodal across text, images, video, and documents. The company claims it rivals Anthropic's Claude Fable 5 on coding and multimodal benchmarks, and it is the first Max-class Qwen model to be open-sourced, with weights promised by end of August. Alibaba shares rose 7% in Hong Kong on the news.

2026-08-04

Alibaba's Qwen-Audio-3.0-ASR-Flash tops OpenAI on new speech benchmark

Alibaba's Qwen team introduced Qwen-Audio-3.0-ASR-Flash, a speech-recognition model with stronger context consistency, domain-term recognition, custom hotwords, and speech-polishing into structured transcripts. Alibaba claims it topped OpenAI on a new speech benchmark, extending China's competitive push in audio AI.

2026-08-02

Alibaba previews 2.4-trillion-parameter Qwen 3.8-Max at WAIC

Alibaba previewed Qwen 3.8-Max at the World AI Conference in Shanghai — a 2.4-trillion-parameter sparse Mixture-of-Experts model that is multimodal across text, vision, video, document, speech and image generation. Alibaba claims it is 'second only to Fable 5,' though public benchmarks are pending, and launched a #QwenGrowthPlan to gather developer feedback on agentic use.

2026-07-30

Alibaba previews Qwen3.8-Max, a 2.4-trillion-parameter multimodal model

Alibaba's Qwen team previewed Qwen3.8-Max-Preview at the World AI Conference in Shanghai — a 2.4-trillion-parameter sparse Mixture-of-Experts multimodal model (text, images, video, documents) with a 1M-token context window. Alibaba claims it is 'second only to Fable 5' among frontier models, available in preview via its Token Plan at 10% of standard pricing. It has now launched a QwenGrowthPlan to gather developer feedback on agentic capabilities.

2026-07-29

Chinese AI models gain US ground as Qwen 3.8-Max, Kimi K3 and DeepSeek advance

Chinese models from Alibaba, Moonshot, Z.ai and DeepSeek are nearing US frontier performance at lower cost and making inroads with US developers. Alibaba previewed its 2.4-trillion-parameter Qwen 3.8-Max, claiming performance 'second only to Fable 5,' while Moonshot's Kimi K3 (2.8T parameters) is set to ship as the largest open-weight model ever, intensifying the open-vs-closed and cost debates.

2026-07-27

Alibaba previews Qwen 3.8 Max, a 2.4T-parameter model it claims trails only Claude

Alibaba previewed Qwen 3.8 Max, a 2.4-trillion-parameter multimodal mixture-of-experts model it claims is 'second only to' Anthropic's Claude across benchmarks, with faster execution in real-world coding tests. The preview is accessible via Alibaba's Token Plan subscription and its Qoder and QoderWork agentic platforms, integrated into Alibaba Cloud.

2026-07-25

Chinese open models Kimi K3 and Qwen rattle Big Tech amid distillation theft accusations

Moonshot AI's Kimi K3 drew so many subscribers it paused signups and was accused of being distilled from Anthropic's Fable, an accusation raised by a former White House science advisor. US open-source lab Arcee argued Chinese models like Kimi K3 and Qwen aren't inherently dangerous despite ban talk. The powerful cheap open models triggered a US tech stock sell-off echoing last year's DeepSeek shock.

2026-07-25

Alibaba previews Qwen 3.8 Max, claims second place behind Claude Fable 5

Alibaba previewed Qwen 3.8 Max at the World AI Conference in Shanghai, claiming 2.4 trillion parameters and a second-place ranking behind only Anthropic's Claude Fable 5. Early reports say the multimodal model clocks faster execution than most competitors in real-world coding tests. The release follows Alibaba's July 6 ban on employees using Claude Code (replaced by its own Qoder tool) and Qwen's approval for Apple Intelligence in China.

2026-07-24

Alibaba previews Qwen 3.8 Max, a 2.4-trillion-parameter model it claims is second only to Fable 5

Alibaba's Qwen team previewed Qwen 3.8 Max, a 2.4-trillion-parameter multimodal model — its first above a trillion parameters — claiming it ranks as the world's second-most powerful AI, trailing only Anthropic's Claude Fable 5. The preview handles text, images, video, and documents via Alibaba's Token Plan and Qoder platforms, with an open-weight release anticipated soon. Alibaba shares rose 5.4% on the news.

2026-07-23

Alibaba previews open-weight Qwen 3.8 Max, claims second globally behind Fable 5 — bans staff from Claude Code

Alibaba previewed Qwen 3.8 Max, a 2.4-trillion-parameter multimodal model going open-weight, claiming it ranks second globally behind only Anthropic's Fable 5. The launch lands alongside Kimi K3 as part of China's open-weights surge. Separately, Alibaba classified Claude Code as high-risk software and banned employee use, citing national-security and distillation-attack concerns.

2026-07-21

Alibaba bans Claude Code and disables Qwen humanoid agents before China regulations

Alibaba banned employees from using Anthropic's Claude Code, labeling it high-risk over alleged back-door vulnerabilities and tracking of Chinese users. It also disabled humanoid AI agent features in Qwen effective July 10, aligning with China's July 15 deadline for anthropomorphic-AI interaction rules, while planning 55+ billion yuan in infrastructure investment.

2026-07-14

Anthropic accuses Alibaba's Qwen lab of massive AI model distillation heist

Anthropic accused Alibaba's Qwen AI lab of orchestrating the largest known AI model theft via a coordinated distillation attack. The operation reportedly used ~25,000 fake accounts generating over 28.8 million Claude interactions between April 22 and June 5, 2026, to extract Claude's coding capabilities for training Qwen. The attack evaded detection by keeping each account below standard rate limits.

2026-07-14

Alibaba releases Qwen-Robot suite, betting on alignment over brute-force scaling

Alibaba released the Qwen-Robot suite on July 12, advancing embodied AI with a focus on model-level alignment rather than brute-force scaling. The suite includes Qwen-RobotNav, which adapts visual attention allocation for mobile control, and Qwen-RobotManip, which standardizes state-action space for manipulation tasks — signaling a shift toward software intelligence competition in robotics.

2026-07-13

Alibaba's Qwen crosses one billion downloads, then pivots to paid API

Alibaba's Qwen model family surpassed Meta's Llama with over one billion cumulative downloads, becoming the world's most widely adopted open-source AI framework, aided by seamless Hugging Face integration and multilingual support. Despite the open-source success, Alibaba is now pivoting its flagship Qwen models toward a closed, API-only paid model.

2026-07-12

Chinese open-source models gain US ground as Beijing weighs export curbs

Open-source Chinese models like Alibaba's Qwen, ByteDance's Doubao, and Z.ai's GLM 5.2 are gaining US adoption at 60–90% lower cost than leading OpenAI and Anthropic models, with GLM 5.2 now a top-five model on some platforms. Meanwhile Beijing is reportedly considering curbing overseas access to China's top AI models, and Chinese tech firms cut roughly 130,000 jobs amid the AI transition.

2026-07-10

Alibaba pivots Qwen to paid, API-only access and sunsets AI agents

Alibaba is moving flagship Qwen models like Qwen3.6-Max and Qwen3.7-Plus from open source to a closed, API-only paid model, monetizing after strong self-hosted adoption. Simultaneously, Alibaba's Qwen and ByteDance's Doubao are discontinuing AI agent features on July 15 amid regulatory pressure.

2026-07-08

Anthropic Alleges Alibaba's Qwen Distilled Knowledge from Claude

The Washington Post reported that researchers found Alibaba's Qwen repeatedly mimicked Claude, misidentifying itself as Claude nearly 100% of the time in intensive tests — evidence Anthropic cites for distillation. Separately, Alibaba reportedly banned employees from using Claude Code, and both ByteDance and Alibaba are disabling humanlike AI custom agents in China ahead of new regulations.

2026-07-07

Alibaba bans Claude Code and disables Doubao/Qwen AI agents for Beijing compliance

Alibaba will bar employees from using Anthropic's Claude Code starting July 10, classifying it high-risk over alleged backdoors and data-retention concerns and directing staff to its own coding tool. Separately, Alibaba's Qwen and ByteDance's Doubao will disable personalized AI-agent functionality on July 15 to comply with China's new interim measures for humanlike interactive AI services, with users given until October 15 to export data before deletion.

2026-07-06

Alibaba bans internal Claude Code usage from July 10, citing an alleged backdoor

Alibaba announced it will bar employees from using Anthropic's Claude Code starting July 10, citing an alleged covert backdoor tied to timezone detection and Chinese cloud keywords. The ban persists despite Anthropic's July 1 restoration of Fable 5 and Mythos 5 with improved safety classifiers.

2026-07-05

Qwen's former lead argues hybrid thinking fell short and makes the case for agents

Junyang Lin, former technical lead of Alibaba's Qwen, reflected on where Qwen3's hybrid thinking modes and dynamic thinking budgets fell short, and argued for a shift from reasoning-thinking to agentic thinking and harder agentic RL infrastructure. His remarks capture an industry pivot from reasoning models to agents.

2026-07-05

Alibaba to Ban Employee Use of Claude Code Over Alleged Backdoor Concerns

Alibaba has banned employees from using Anthropic's Claude Code at work after the tool drew scrutiny for features that can help identify China-linked users, per Reuters. Employees are being told to use Alibaba's own Qoder platform, escalating a feud in which Anthropic accused Alibaba of illicitly extracting Claude's model capabilities.

2026-07-04

Cheaper Chinese and open-source models reshape enterprise AI budgets

Soaring token bills are pushing businesses toward cheaper open-source and Chinese models, with the four most popular models on OpenRouter all Chinese and DeepSeek in the top spot at as little as 18 cents per million tokens. Coinbase is experimenting with GLM 5.2 and Kimi 2.7 defaults to cut costs, with Uber reportedly burning its 2026 AI budget in four months.

2026-07-02

Anthropic accuses Alibaba of largest known distillation attack on Claude

Anthropic publicly accused Alibaba of conducting the largest known distillation attack against its Claude models, an incident in which capabilities were allegedly extracted without authorization. The disclosure highlights escalating competition between Western and Chinese labs and vulnerabilities in frontier-model deployment.

2026-07-02

Alibaba shares hit 16-month low after Anthropic's IP-extraction accusation

Alibaba slid to a 16-month low after Anthropic accused the company of 'illicitly' accessing and extracting capabilities from its Claude AI models. The accusation reignited debate over model distillation, IP theft and the ethics of training on competitor outputs, landing in the same week Anthropic navigated US export restrictions on its own frontier models.

2026-06-28

Alibaba's Qwen-AgentWorld beats seven agent benchmarks by predicting environments

Alibaba's Qwen team released Qwen-AgentWorld, two models trained to predict agent environments rather than act within them. This 'world modeling' approach outperformed on seven agent benchmarks — including three not seen during training — across software engineering, search and Android. The Mixture-of-Experts models support 256K context, were trained on 10M+ environment-interaction trajectories, and reportedly exceed gains from traditional real-environment reinforcement learning.

2026-06-26

Alibaba unveils Qwen Robot Suite to challenge NVIDIA in physical AI

Alibaba's Tongyi Lab unveiled the Qwen Robot Suite, a family of embodied AI models — Qwen-RobotNav for navigation, Qwen-RobotWorld as a video 'world model,' and Qwen-RobotManip for physical execution — pushing AI beyond chatbots into robotics and challenging NVIDIA's physical-AI ambitions.

2026-06-25

Alibaba unveils Qwen Robot Suite of embodied AI models to challenge NVIDIA in physical AI

Alibaba's Tongyi Lab unveiled the Qwen Robot Suite, a family of embodied AI models enabling robots to perceive, reason, and act in physical environments. It includes Qwen-RobotNav for navigation, Qwen-RobotWorld as a video 'world model,' and Qwen-RobotManip for physical execution, marking Alibaba's push beyond chatbots into robotics and physical AI in direct competition with NVIDIA.

2026-06-24

Alibaba unveils Qwen Robot Suite, its first AI models for physical robots

Alibaba unveiled the Qwen Robot Suite, its first set of AI models built specifically for robots, marking a shift toward embodied AI. Developed by Tongyi Lab, it includes Qwen-RobotNav for navigation, Qwen-RobotManip for manipulation, and Qwen-RobotWorld for predicting future physical states. The models are in pilot testing with selected Alibaba Cloud enterprise clients.

2026-06-21

Alibaba launches Qwen Robot Suite for embodied AI

Alibaba unveiled its Qwen Robot Suite, introducing three AI models — Qwen-RobotNav, Qwen-RobotManip, and Qwen-RobotWorld — designed to help machines navigate, manipulate objects, and predict physical states. The launch marks Alibaba's push into embodied AI, with pilot testing already underway with select Alibaba Cloud enterprise clients.

2026-06-20

Alibaba forms Token Foundry unit and plans to exceed ¥380B AI investment

Alibaba established Token Foundry as a new business unit, tightening its AI organization after creating Alibaba Token Hub, and said it plans to exceed its original ¥380 billion (~$56B) three-year AI investment commitment, stating margin performance is secondary to AI leadership. It also released a new Qwen model it claims surpasses DeepSeek-V3.

2026-06-16

Alibaba Cloud launches QwenCloud and Qwen3.7-Max at first international Qwen Conference

Alibaba Cloud held its inaugural international Qwen Conference 2026, launching QwenCloud and the new flagship Qwen3.7-Max model, optimized for agentic tasks with a 1-million-token context window. The company is positioning itself as a full-stack AI contender aiming to expand overseas with infrastructure, open-source models, and cloud services.

2026-06-12

Alibaba's cloud division begins AI-driven 'quiet' layoffs as Beijing pushes adoption

An engineer at Alibaba's cloud division said AI-driven headcount reductions have begun, unfolding through gradual cuts and attrition rather than a single mass round. The trend reflects broader 'quiet' layoffs across Chinese tech firms as Beijing promotes AI adoption while avoiding visible job losses that threaten social stability.

2026-06-11

Alibaba launches Qwen3.7-Plus, a multimodal agent model that sees screens and writes code

Alibaba's Qwen team released Qwen3.7-Plus, a multimodal agent model capable of visual perception, GUI control and code generation within an autonomous agent loop. Available via API on Alibaba Cloud's Bailian platform, it processes text, images and video to operate software, leads Alibaba's GUI-grounding benchmarks, and is priced aggressively at $0.40/$1.60 per million input/output tokens.

2026-06-10

Alibaba launches Qwen3.7-Plus, a multimodal 'computer-use' agent

Alibaba's Qwen team released Qwen3.7-Plus, a multimodal agent model combining visual perception, GUI control and code generation in an autonomous agent loop. Available via API on Alibaba Cloud's Bailian platform, it accepts text, images and video to read screens, navigate applications, write code from visual templates and invoke tools without human intervention. Alibaba says it leads its GUI-grounding benchmarks.

2026-06-08

Alibaba launches Qwen3.7-Plus multimodal agent that reads screens and writes code

Alibaba's Qwen team released Qwen3.7-Plus on June 2, a multimodal agent model combining visual perception, GUI control and code generation in an autonomous agent loop, generally available via Alibaba Cloud's Bailian platform. It takes text, images and video to read screens, navigate apps, generate code from visual templates and invoke external tools autonomously, as Alibaba opens Qwen to third parties for agent-powered commerce.

2026-06-07

Alibaba's open-weight Wan 2.7 leads Chinese video-AI surge as OpenAI retreats

A Forbes analysis details how Chinese labs are winning the video-AI race OpenAI largely abandoned, led by Alibaba's open-weight Wan 2.7 running a Llama-style open-source flanking strategy. The piece frames open weights as a distribution lever, while CNBC reports China may adopt US-style talent poaching as labs compete for AI researchers.

2026-06-07

Alibaba's Qwen3.7-Max claims smartest-Chinese-LLM crown with the lowest frontier hallucination rate

Alibaba released Qwen3.7-Max, its latest proprietary LLM, positioning it as the smartest Chinese model and third-fastest overall on the Artificial Analysis Intelligence Index. Built for long-running agentic work, coding and scientific discovery, it posted a 23% hallucination rate — the lowest among frontier models tested — though partly by declining to answer over half the prompts.

2026-06-06

Alibaba's Qwen team launches Qwen3.7-Plus with vision and agentic features

Alibaba's Qwen team released Qwen3.7-Plus, a multimodal large language model on Alibaba Cloud's Bailian platform that understands image and video inputs and adds agentic capabilities including deep reasoning, self-programming, tool invocation and autonomous iteration. It builds on the prior Qwen3.7 generation by adding multimodal support.

2026-06-05

Alibaba's Qwen3.7-Plus adds vision and five agentic skills

Alibaba's Qwen team released Qwen3.7-Plus, a multimodal model on its Bailian platform adding image and video understanding plus five agentic skills: deep reasoning, self-programming, tool invocation, output verification and autonomous iteration. Qwen3.7-Plus-Preview ranked 16th on Vision Arena.

2026-06-04

Alibaba releases Qwen 3.7 Plus, a low-cost multimodal GUI agent

Alibaba's Qwen team released Qwen 3.7 Plus, a multimodal model with vision input adding deep reasoning, self-programming, tool invocation and autonomous iteration, positioned as a roughly $0.40 GUI agent for high-volume enterprise workloads. It ranks 16th globally on the Vision Arena leaderboard.

2026-06-03

Alibaba's Qwen3.7-Max ranks fourth on Code Arena, topping deployed OpenAI and Google models

Alibaba's Qwen3.7-Max placed fourth on Code Arena's WebDev leaderboard for AI coding, surpassing deployed models from OpenAI and Google. The model — boasting over one trillion parameters and a one-million-token context window — excels at building web applications and agent-driven workflows including coding, office automation, and complex long-running tasks.

2026-06-02

Alibaba's Qwen3.7-Max cracks top of Code Arena WebDev leaderboard, beating OpenAI and Google

Alibaba's Qwen3.7-Max ranked fourth on Code Arena's WebDev leaderboard, surpassing deployed OpenAI and Google models and becoming the only non-US developer in the top five. Built for agent-driven workflows, it's claimed to run autonomously up to 35 hours without performance degradation, and also hit #3 on the IT-task ITbench-AA benchmark.

2026-05-31

Alibaba Qwen3.7-Max ranks #4 globally on Code Arena WebDev, beating OpenAI and Google

Alibaba's Qwen3.7-Max reached #4 on Code Arena's WebDev leaderboard, the only non-US developer in the top five (behind several Claude models) and beating deployed OpenAI and Google models on web application building. Alibaba also reported Qwen3.5 hitting 580 tps on the TokenSpeed engine.

2026-05-30

Qwen3.7-Max Hits #4 on Code Arena, On Par With Claude Opus 4.6

Alibaba's Qwen3.7-Max debuted at #4 on Code Arena, level with Claude Opus 4.6, making it the top-ranked Chinese lab on the board. Alibaba says more is shipping.

2026-05-29

Alibaba Cloud unveils Qwen3.7-Max and Qwen Cloud AI-native platform

Alibaba Cloud launched Qwen3.7-Max on Model Studio with a 1M-token context window and a 56.6 score on the Artificial Analysis Intelligence Index v4.0 — #5 overall and the highest Chinese model. Pricing is roughly half of Claude Opus 4.7. Qwen Cloud, a new AI-native platform, was unveiled alongside for global enterprise agent workloads.

2026-05-28

Alibaba Cloud debuts Qwen3.7-Max and global agentic AI ecosystem at first international Qwen Conference

Alibaba Cloud unveiled Qwen3.7-Max — ranked fifth globally and first among Chinese models on Artificial Analysis's Intelligence Index — alongside infrastructure upgrades, an AI-native platform and a new agent ecosystem at its first international Qwen Conference. Implicit caching is also live on Qwen3.7-Max, automatically reducing cost and latency.

2026-05-27

Alibaba unveils Qwen 3.7 Max with 1M-token context, plus Zhenwu M890 chip and Panjiu AL128 supernode

At the 2026 Alibaba Cloud Summit, Alibaba officially unveiled Qwen 3.7 Max with a 1M-token context window, scoring 56.6 on the Artificial Analysis Intelligence Index v4.0 — #5 overall and highest among Chinese models. T-Head also unveiled the Zhenwu M890 AI chip and Panjiu AL128 supernode server, completing a vertical stack from silicon to flagship LLM. Pricing is roughly half of Claude Opus 4.7.

2026-05-26

Alibaba's Qwen3.7-Max ran autonomously for 35 hours optimizing a kernel on a never-before-seen custom chip

Alibaba's Qwen team released Qwen3.7-Max, a proprietary agent-first model with a 1M-token context window, available exclusively via API. In a real-world test it autonomously optimized an attention kernel for 35 straight hours on a custom chip architecture it had never seen before — a striking demonstration of long-horizon agent capability.

2026-05-25

Alibaba and China internet giants plan $84B 2027 AI capex as 'Seven Titans' stocks slump

Alibaba and other Chinese internet giants are eyeing $84B in combined AI investment for 2027, even as China's 'Seven Titans' tech stocks slump under deflationary pressure. The bet sits in stark contrast to share-price weakness and signals that Chinese AI capex is being treated as strategic rather than market-driven.

2026-05-25

Qwen3.7-Max runs autonomously for 35 hours to optimize code on its own custom chip

Alibaba's Qwen team released Qwen3.7-Max, a proprietary agent-foundation model that ran autonomously for 35 hours doing kernel optimization on a T-Head ZW-M890 PPU. It's exclusively available via Alibaba Cloud Model Studio API with a 1M-token context window.

2026-05-24

Qwen 3.7-Max: 35-hour autonomous agent runs, paid-API shift away from open frontier

Alibaba's Qwen Team released Qwen 3.7-Max, a proprietary LLM capable of autonomous agentic work for up to 35 hours without performance degradation, with support for external harnesses including Anthropic's Claude Code. The release marks a strategic shift: Alibaba's most powerful frontier models are now offered via paid APIs and subscriptions rather than fully open-source weights, putting Qwen squarely against GPT-5, Claude and Gemini.

2026-05-23

Qwen3.7-Max debuts with 35-hour autonomous execution and 56.6 on Artificial Analysis index

Alibaba's Qwen team released Qwen3.7-Max, a proprietary model built for the 'agent era' that demonstrated ~35 hours of continuous autonomous execution in internal tests. It scored 56.6 on the Artificial Analysis Intelligence Index — a 4.8-point jump over Qwen3.6-Max-Preview — and offers a 1M-token context window with sharper scientific reasoning, stronger agentic capabilities, better coding, and reduced hallucination.

2026-05-22

Alibaba opens all of Taobao to Qwen, launching agentic shopping across 4B+ products

Alibaba has fully integrated its Qwen App with Taobao's catalog of 4 billion+ products, enabling conversational shopping inside the Taobao app. The Qwen-powered assistant lets users browse, compare, order, and manage deliveries via natural language. Features include AI Virtual Try-On and automated discount aggregation, with Qwen accessing a 'skills library' for logistics and after-sales service.

2026-05-19

Alibaba's Qwen3.6-35B-A3B hits 24.6% on Terminal-Bench 2.0 with just 3B active params

Alibaba's Qwen team released Qwen3.6-35B-A3B — a sparse MoE model with 35B total parameters but only 3B active per token — that hit 24.6% on the public Terminal-Bench 2.0 leaderboard. The Qwen team also released Qwen3.5-9B alongside it. The 35B/3B model is being highlighted as a budget-friendly choice for developers building AI products on tight inference economics.

2026-05-18

Alibaba Cloud revenue jumps 38% YoY to $6B; new Qwen3.6-35B-A3B MoE hits Terminal-Bench 2.0 with 3B active params

Alibaba's fiscal Q4 2026 Cloud Intelligence Group revenue hit $6.035B, up 38% YoY, with AI-related products posting an 11th straight quarter of triple-digit growth and contributing 30% of external cloud revenue. The Qwen team simultaneously released Qwen3.6-35B-A3B and Qwen3.5-9B MoE models, scoring 24.6% on Terminal-Bench 2.0 while activating only 3B parameters per token.

2026-05-17

Alibaba Cloud revenue up 38%, AI products hit 11th straight triple-digit quarter

Alibaba reported Cloud Intelligence Group revenue of $6.035B (+38% YoY), with AI products at RMB 8.97B contributing 30% of external cloud revenue and posting an 11th consecutive triple-digit quarter. CEO Eddie Wu said full-stack AI has moved 'from incubation to commercialization at scale.' New Qwen3.6-35B-A3B and Qwen3.5-9B MoE models posted strong Terminal-Bench 2.0 results.

2026-05-17

Alibaba launches Qwen-Image-2.0 with doubled compression and 10x faster generation

Alibaba released Qwen-Image-2.0, an updated image generation model that doubles compression and slashes generation steps from 40 to 4. Architectural changes include a reworked VAE for harder compression and a dedicated prompt-expansion module, aimed at faster and cheaper training and inference.

2026-05-16

Qwen team ships Qwen-Image-VAE-2.0 with high-compression VAEs

Alibaba's Qwen team released Qwen-Image-VAE-2.0, a suite of high-compression Variational Autoencoders aimed at improving generative image quality by fixing the compression layer. The technical report appeared on arXiv May 13 and Hugging Face Papers May 14, detailing Global Skip Connections and semantic alignment.

2026-05-15

Alibaba AI/cloud revenue jumps 38% as Qwen spend pressures profit; Qwen-Image-VAE-2.0 ships

Alibaba posted a 38% jump in AI/cloud revenue with AI products now 30% of external cloud customer revenue. Bloomberg Intelligence notes 90%+ of March-quarter China e-commerce profit was redeployed into Qwen user acquisition. Qwen-Image-VAE-2.0 also released, fixing the compression layer behind image generation.

2026-05-15

Alibaba merges Qwen AI into Taobao for conversational shopping at scale

Alibaba is integrating Qwen with Taobao and Tmall to replace keyword search with AI-driven conversational shopping. Users can browse, compare, and buy across four billion-plus products by interacting with a Qwen-powered assistant inside the Qwen app, plus a skills library for logistics tracking and personalized recommendations.

2026-05-14

Alibaba research: ReVision cuts GUI agent costs; Agent-BRACE and anchored bipolicy self-play

Alibaba researchers published three notable agent-systems papers: ReVision exploits temporal visual redundancy across GUI screenshots to scale computer-use agents under fixed context budgets; Agent-BRACE decouples beliefs from actions in long-horizon partially-observable tasks; and an 'anchored bipolicy' self-play method exposes safety self-consistency gaps that standard red-teaming misses.

2026-05-13

Qwen merges into Taobao; Qwen3-Max-Thinking pushes 'A2A' agent commerce

Alibaba is preparing to deeply integrate Qwen AI into Taobao and Tmall, turning keyword shopping into conversational, agent-driven browsing across 4 billion+ products with a logistics and after-sales 'skills library.' It also introduced Qwen3-Max-Thinking, a test-time-scaled reasoning model with native tool use, and expanded its Accio Work autonomous enterprise AI agent — pitching a shift from B2B to 'A2A' (agent-to-agent) trade.

2026-05-13

Alibaba projects $4.4B annualized AI revenue by year-end

CEO Eddie Wu said AI model and application services revenue will top 10B yuan ($1.47B) annualized this quarter and reach 30B yuan ($4.4B) by year-end — Alibaba's first concrete AI monetization disclosure. Separately, ex-Qwen lead Junyang Lin is raising at a $2B valuation for a new AI lab.

2026-05-13

Alibaba fuses Qwen with Taobao for agentic shopping across 4B-product catalog

Alibaba is preparing a major integration of its Qwen AI app with Taobao and Tmall, giving the assistant access to over 4 billion products plus a 'skills library' for logistics, customer service, and personalized recommendations based on order history. Consumers will browse, compare, and purchase by chatting with the AI rather than navigating product listings. Inside Taobao itself, a Qwen-powered shopping assistant will offer virtual try-ons and 30-day price tracking.

2026-05-12

Qwen AI Glasses S1 gets proactive AI + spatial 3D display upgrade

Alibaba rolled out a major software upgrade for its Qwen AI Glasses S1, adding proactive AI capabilities (contextual suggestions and reminders without explicit voice commands) and spatial 3D display technology, plus several new lifestyle-focused AI services.

2026-05-09

Alibaba's Qwen 3.6 Max reportedly outperforms Claude 4.5 Opus on instruction-following and agentic coding

Alibaba released Qwen 3.6 Max, which independent testers say outperforms Claude 4.5 Opus and Zhipu's GLM 5.1 on instruction following, agentic coding and multimodal processing, with sharper visual reasoning and document analysis. The release lands alongside Moonshot AI's $2B raise at a $20B valuation as the Chinese open-weights ecosystem closes the gap with US frontier labs.

2026-05-08

Morgan Stanley names Alibaba top China AI pick; 41% of CIOs choose Qwen

A Morgan Stanley CIO survey found 41% of Chinese CIOs picked Alibaba for AI deployment help — up from prior surveys — and 30% expect Alibaba to capture the largest incremental share of AI spending this year. AI now represents 43% of the Hang Seng Tech pipeline.

2026-05-04

Qwen releases Qwen-Scope, open suite of sparse autoencoders

Alibaba's Qwen team released Qwen-Scope, an open suite of sparse autoencoders for the Qwen model family, enabling steering of model outputs by manipulating internal features, classification with minimal examples, and tracing code-switching patterns. The release pushes the interpretability frontier further into open weights.

2026-05-02

Alibaba Cloud holds 37% China share as Q4'25 cloud spend rises 26%, fueled by Qwen3.5

Per Omdia, Mainland China cloud infrastructure spending rose 26% YoY in Q4 2025, driven by AI and agent growth. Alibaba Cloud retained 37% market share with triple-digit AI revenue growth for the tenth consecutive quarter, bolstered by Qwen3.5 and Model Studio (Bailian) enhancements.

2026-04-29

Qwen passes 1B downloads as VW and BYD integrate it into vehicles

Alibaba announced Qwen has surpassed 1 billion cumulative downloads and supported the creation of 200,000+ derivative models, cementing it as a major global open-weight platform. Qwen is being integrated into cars from Volkswagen, BYD, Geely, Li Auto and SAIC Volkswagen, running on NVIDIA's in-car chip system to enable voice-driven hotel bookings, food delivery and package tracking.

2026-04-28

Alibaba puts Qwen voice AI in BYD, Geely, VW China vehicles

Alibaba is integrating its Qwen AI model into vehicles from BYD, Geely, and a local VW joint venture, enabling drivers to use voice commands to track packages, order food, and reserve hotels. The system blends on-device processing with cloud compute for multi-step service-linked tasks. Separately, Alibaba research showed component-type-aware LoRA placement improves hybrid Qwen3.5 fine-tuning over uniform placement.

2026-04-27

Qwen3.6-27B beats 15x-larger predecessor on coding; Qwen lands in BYD/Volkswagen EVs

Alibaba's open-source Qwen3.6-27B outperforms its 15x-larger predecessor on most coding benchmarks with just 27B parameters, continuing the small-dense-model trend. Qwen is also being integrated into BYD and SAIC Volkswagen EVs running on Nvidia automotive chips, enabling in-car voice booking for China Eastern flights and food delivery.

2026-04-26

Alibaba's Qwen AI enters electric vehicles with BYD and SAIC Volkswagen, enables in-car voice bookings on Nvidia chips

Alibaba announced Qwen AI model integration into vehicles from BYD and SAIC Volkswagen (a Volkswagen joint venture), enabling in-car voice commands for services like food delivery, hotel bookings, and China Eastern Airlines flight reservations. Qwen will operate on Nvidia's automotive chip platform. Alibaba separately released Qwen3.6-27B, a dense 27B-parameter open-weight model that outperforms 397B MoE models on agentic coding benchmarks.

2026-04-25

Qwen AI agent books flights directly with China Eastern Airlines

Alibaba integrated its Qwen AI agent with China Eastern Airlines for flight bookings via natural-language chat, a significant commercial deployment of agentic AI facilitating real-world transactions beyond traditional app flows.

2026-04-24

Alibaba Releases Qwen3 Series with Hybrid Thinking and 235B MoE Model

Alibaba's Qwen team released the Qwen3 series, including dense models from 0.6B to 32B parameters and a 235B Mixture-of-Experts flagship, all available as open weights. Qwen3 introduces hybrid thinking mode allowing models to switch between fast response and extended chain-of-thought reasoning. The 32B dense model claims to outperform DeepSeek-R1 and o1 on math and coding benchmarks while supporting 119 languages.

2026-04-23

Alibaba's Qwen Partners with China Eastern Airlines for Agent Experience

Alibaba Group's Qwen App initiated its first external AI agent partnership with China Eastern Airlines, enabling users to manage entire flight booking and check-in processes within a single chat interface. This collaboration aims to streamline travel experiences by eliminating the need for multiple platforms, with Alibaba planning to expand Qwen's agentic capabilities through additional partnerships.

2026-04-23

Alibaba Qwen Team Releases Qwen3.6-27B Dense Model for Agentic Coding

Alibaba's Qwen Team released Qwen3.6-27B, a 27-billion-parameter dense model specifically optimized for coding agents. The model features significant improvements in agentic coding capabilities, a novel Thinking Preservation mechanism, and hybrid architecture blending Gated DeltaNet linear attention with traditional self-attention. It reportedly outperforms much larger 397B MoE models on agentic coding benchmarks.

2026-04-23

Alibaba Releases Qwen3.6-Max Preview with Enhanced Coding and Knowledge Capabilities

Alibaba's Qwen team released Qwen3.6-Max-Preview, featuring significantly improved agent programming abilities, stronger world knowledge, and better instruction following compared to Qwen3.6-Plus. The preview model shows notable gains on benchmarks including SkillsBench (+9.9), SciCode (+10.8), and Terminal-Bench 2.0 (+3.8), positioning it competitively against leading frontier models.

2026-04-22

Alibaba releases Qwen3.6-35B-A3B model demonstrating competitive performance against Google Gemini at fraction of cost

Alibaba released Qwen3.6-35B-A3B on April 16, 2026, featuring 35 billion total parameters with 3 billion active during inference. The model demonstrates competitive agentic coding performance against larger models like Google's Gemini 3.1 Pro while offering significant cost advantages at $0.38 per million input tokens compared to Gemini's $2.00. It also shows strong multimodal perception and reasoning abilities.

2026-04-21

Qwen3.6-35B-A3B Demonstrates Competitive Performance Against Google Gemini

Alibaba released Qwen3.6-35B-A3B on April 16, 2026, with 35 billion total parameters and 3 billion active during inference. The model demonstrates competitive agentic coding performance against larger models like Google's Gemini 3.1 Pro while offering significant cost advantages at $0.38 per million input tokens.

2026-04-20

Alibaba's Qwen 3.6 35B Model Runs Locally on Apple M5 Max MacBooks with Cloud-Comparable Performance

Developers are running Alibaba's Qwen 3.6 35B model locally on Apple's M5 Max MacBook Pro with 128GB unified memory, achieving inference quality comparable to cloud-based models like Claude and GPT-4. The development challenges traditional cloud AI business models by demonstrating capable local inference on consumer hardware.

2026-04-19

Alibaba restructures AI leadership with Qwen3.6-Plus topping leaderboards as 397B-parameter multimodal model

Alibaba Group formed a new AI-focused technology committee led by CEO Eddie Wu and appointed Li Feifei as Cloud CTO to sharpen AI strategy. The company launched Qwen3.6-Plus, which topped multiple AI leaderboards, following the February 2026 release of Qwen3.5—a 397-billion-parameter multimodal open-weight model with native agentic capabilities supporting over 200 languages.

2026-04-18

Alibaba's Qwen Models Achieve Nearly 1 Billion Downloads, Dominating Open-Source AI Market

Alibaba's Qwen AI models reportedly captured over 50% of the global open-source AI market, reaching 942.1 million cumulative downloads by March 2026. The latest Qwen 3.5 release achieves performance comparable to leading closed-source models while offering significantly lower token costs, with over 100,000 derivative models built on the platform.

2026-04-15

Alibaba Releases Qwen3 Model Family Including 235B MoE Flagship

Alibaba's Qwen team open-sourced the Qwen3 family, ranging from Qwen3-0.6B to Qwen3-235B-A22B (a mixture-of-experts model activating 22B parameters per forward pass). The family supports 119 languages, features a hybrid thinking/non-thinking mode switchable via a chat tag, and is released under Apache 2.0. Benchmark results show Qwen3-235B-A22B matching or exceeding GPT-4.1 and Claude 3.7 Sonnet on several coding and math tasks. Alibaba Cloud captured more than 50% of global open-source model downloads as of March, with Qwen reaching nearly 1 billion cumulative downloads.

2026-04-14

Alibaba's Qwen Captures Over 50% of Global Open-Source AI Model Downloads

Alibaba's Qwen AI model family dominates over half of the global open-source AI market with 942.1 million cumulative downloads by March 2026, according to Interconnect AI research. In February alone, Qwen recorded 153.6 million downloads—more than double the combined downloads of the next eight leading competitors including Meta, DeepSeek, and OpenAI on Hugging Face. The Qwen 3.5 model, open-sourced in February, delivers performance comparable to leading closed-source models from OpenAI and Anthropic but at one-tenth the token usage cost of Google Gemini. Over 100,000 derivative models based on Qwen have been developed, demonstrating widespread adoption in the global developer ecosystem.

2026-04-13

More vendors

AnthropicOpenAIGoogleAWSAzureMetaxAINVIDIAMistralAppleHugging FaceDeepSeekSamsung

← Browse all AI stories