Aakib Ansari.
Back to articles
News Brief

Alibaba's Qwen 3.8 Max Puts 2.4 Trillion Parameters on a Public API and Promises Open Weights by August 10

Md Aakib Ansari
Md Aakib AnsariWeb Developer & AI Tools Reviewer
2 min readModel: Qwen 3.8 Max
Alibaba's Qwen 3.8 Max Puts 2.4 Trillion Parameters on a Public API and Promises Open Weights by August 10

Alibaba launched Qwen 3.8 Max on August 3, 2026. As the flagship release of the new Qwen 3.8 series, it is their largest model to date: 2.4 trillion total parameters, 95 billion active per token, using a sparse Mixture-of-Experts architecture. The release brings the highly anticipated qwen open weights roadmap to the forefront, offering a model that accepts text, image, video, and document inputs in a single request and supports a 1-million-token context window. It is live now via Alibaba Cloud Model Studio and the QwenWork platform.

API pricing is set at 12 RMB per million input tokens and 36 RMB per million output tokens — roughly comparable to mid-tier US frontier model pricing on a purchasing-power-adjusted basis. The bigger near-term story is on the open-weights side: Alibaba announced that Qwen 3.8 Max will be the first "Max-class" model in the Qwen family to have weights released publicly, with that release scheduled for the week of August 10. That would make it one of the largest open-weight models available (and the largest open weight LLM to date), exceeding Kimi K3's 1.4-trillion-parameter weight release from late July.

Alibaba's internal evaluations position the model as second only to Claude Fable 5 on benchmark composites, outperforming [GPT-5.6 Sol](/models/gpt-5-6-sol) and Gemini 3.1 Pro on coding and multimodal understanding tasks, while trailing in some general-purpose reasoning categories — a pattern consistent with how flagship chinese open source llm variants have typically landed relative to western frontier labs. Independently, Qwen 3.8 Max has already appeared on the LMSYS Chatbot Arena leaderboard, where it ranks first among Chinese models on text tasks and second globally in the multimodal/visual analysis category — the only third-party signal available at launch.

The release lands a week after Kimi K3's open weights went public, and two days after the EU AI Act's GPAI transparency rules took effect — a regulatory backdrop that may shape how Alibaba positions open-weight access in European markets. This release highlights Alibaba's open source AI roadmap, cementing their position in the open weights ecosystem. For context on where the broader Qwen family sits, see our earlier coverage of Qwen 3.7 Flash, the cost-tier counterpart that launched at $0.03 per million tokens in July. Qwen 3.8 Max is now clearly the flagship generation; the 3.7 series is the volume tier.

Frequently Asked Questions

What is Qwen 3.8 Max and who made it?
Qwen 3.8 Max is Alibaba's flagship large language model as of August 2026. It uses a sparse Mixture-of-Experts architecture with 2.4 trillion total parameters and 95 billion active parameters per forward pass. It is multimodal, accepting text, image, video, and document inputs, with a 1-million-token context window.
When will Qwen 3.8 Max open weights be released?
Alibaba announced that open weights for Qwen 3.8 Max will be released during the week of August 10, 2026. This would make it the first 'Max-class' Qwen model to have publicly available weights and one of the largest open-weight models ever released.
How does Qwen 3.8 Max compare to GPT-5.6 and Claude Fable 5?
According to Alibaba's internal evaluations, Qwen 3.8 Max ranks second only to Claude Fable 5 on frontier benchmark composites, outperforming GPT-5.6 Sol on coding and multimodal tasks. It trails on some general-purpose reasoning benchmarks. Third-party validation from LMSYS Chatbot Arena shows it as the top Chinese model on text tasks and second globally for multimodal analysis.
Where can I access Qwen 3.8 Max?
Qwen 3.8 Max is available now via Alibaba Cloud Model Studio API and the QwenWork platform. API pricing is 12 RMB per million input tokens and 36 RMB per million output tokens. Open weights are expected by the week of August 10, 2026.

Related Articles

Meta Releases Muse Glimmer: Apache 2.0 Licensed 30B Local Agent Model
News Brief2 min read
Meta Releases Muse Glimmer: Apache 2.0 Licensed 30B Local Agent Model

Meta Superintelligence Labs has released Muse Glimmer, a 30-billion parameter open-weight model distilled from its proprietary Muse Spark flagship. Published under an Apache 2.0 license, Glimmer is purpose-built for offline, on-device agentic workloads like coding, debugging, and file management on consumer hardware.

Ant Group's inclusionAI Team Releases Ling 3.0 Flash FP8 Under MIT License
News Brief2 min read
Ant Group's inclusionAI Team Releases Ling 3.0 Flash FP8 Under MIT License

Ant Group's inclusionAI team has released Ling 3.0 Flash FP8, a highly efficient 124-billion parameter Mixture-of-Experts (MoE) model. Featuring an MIT license and a custom hybrid attention architecture, the model reduces active parameters to 5.1 billion per token, matching the performance of much larger models while dramatically lowering operational costs.

Liquid AI's LFM2.5-2.6B Matches Models Three Times Its Size on Agentic Tasks
News Brief2 min read
Liquid AI's LFM2.5-2.6B Matches Models Three Times Its Size on Agentic Tasks

Liquid AI released LFM2.5-2.6B on August 4, a 2.69B-parameter on-device model purpose-built for agentic workloads. Using a hybrid architecture of short convolution blocks and grouped query attention, it runs under 2.5 GB of memory and reaches approximately 220 tokens/s on Apple M5 Max — while matching or exceeding Qwen3.5-9B on tool use and instruction-following benchmarks.