Alibaba's Qwen3.8 Max is a 2.4 trillion parameter sparse MoE model (95B active) with a 1M token context window, native text/image/video input, and API pricing of $2/$6 per million tokens.

Create a landscape editorial hero image for this Studio Global article: What are the key details of Alibaba's Qwen3.8-Max model launch, including its mixture-of-experts architecture, parameter count, context wind. Article summary: Here are the key details of Alibaba's Qwen3.8-Max launch across all the dimensions you asked about.. Topic tags: general, general web, user generated, news. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts with fake numbers, clickbait thumbnails, icons, and tiny thumbnail layouts. Make it useful as an illustrative visual, not
On August 3, 2026, Alibaba released Qwen3.8-Max, its largest and most capable AI model to date — a 2.4-trillion-parameter sparse mixture-of-experts (MoE) model that pushes the boundaries of what's possible with efficient inference, multimodal input, and agentic workflows. But the headline numbers tell only part of the story: the launch simultaneously introduced a revenue-sharing model for large commercial users of the open weights, an enterprise AI agent platform called QwenWork, and a new developer platform that opens Alibaba's AI ecosystem to third parties for the first time.
Qwen3.8-Max is built on a sparse MoE architecture, the first Max-class Qwen model to use this design . With 2.4 trillion total parameters, only about 95 billion parameters are activated per token — roughly 4% of the total — which dramatically reduces inference costs compared to dense models of similar scale
.
The model supports a 1 million token context window (maximum practical input is ~991K tokens, dropping to ~983K with thinking mode enabled) and maximum output of 131K tokens .
For the first time in a Qwen model above 1 trillion parameters, Qwen3.8-Max offers native multimodal input: text, images, and video . Output is text-only.
On the independent Arena leaderboards, Qwen3.8-Max ranks:
It is described as the top Chinese text model on Arena and second globally for multimodal/visual performance . On PaperBench, it reports the highest score at 93.0
. On Terminal-Bench 2.1, it scores 86.6, ahead of Claude Opus 4.8 and Claude Fable 5 (both at 84.6) but behind GPT-5.6 Sol (max) at 88.8
.
The API launched immediately through QwenCloud and Alibaba Cloud Model Studio at :
Rate limits are set at 2 million tokens per minute and 15,000 requests per minute .
Open weights for Qwen3.8-Max and the smaller Qwen3.8-27B were expected around August 10, 2026 — making this the first open-weight release for a Max-class Qwen model .
In a breaking story reported by Reuters on August 7, 2026, Alibaba plans to require large commercial users of Qwen3.8-Max's open weights to negotiate a revenue-sharing agreement . The move follows the licensing playbook set by Moonshot AI's Kimi K3, which seeks revenue share of up to 30% from resellers generating more than $20 million in annual sales
.
Key details (based on Reuters reporting and multiple confirmatory outlets):
This marks a significant shift from earlier Qwen releases, which offered completely unrestricted commercial use . If implemented as reported, it would make Alibaba the second major Chinese AI lab to attach commercial terms to open-weight releases within a month
.
Simultaneously with Qwen3.8-Max, Alibaba launched QwenWork (Chinese name: 千问办公), an all-in-one workplace AI agent platform .
QwenWork is designed for tasks such as document preparation, application development, multimodal generation, web development, browser automation, and scheduled tasks — all powered by the new flagship model .
On August 10, 2026, Alibaba launched the Qianwen Open Platform, giving ecosystem partners and third-party developers access to build custom AI-agent applications integrated across smartphones, PCs, and AI glasses . This was the first time Alibaba opened its previously internal Qianwen AI platform to outside developers
.
The first group of partners spans more than 10 sectors including logistics, real estate, and wealth management, with participants such as SF Express, Ziroom, and Midea Group .
Alibaba's August 2026 launch is more than a model release. It's a coordinated strategy that combines:
The revenue-sharing component is the most consequential for the AI industry: if adopted broadly, it could redefine what "open-source AI" means in practice — especially for large-scale commercial deployments.
Studio Global AI
Use this topic as a starting point for a fresh source-backed answer, then compare citations before you share it.
Alibaba's Qwen3.8 Max is a 2.4 trillion parameter sparse MoE model (95B active) with a 1M token context window, native text/image/video input, and API pricing of $2/$6 per million tokens.