Briefly
- Qwen3.8-Max runs 2.4 trillion parameters and is the primary Max-class Qwen mannequin Alibaba will launch as open weights.
- It ships with setup directions for Anthropic’s Claude Code and OpenAI’s Codex—rival instruments.
- On Alibaba’s personal benchmark desk, Fable 5 wins 15 of 31 exams. Qwen3.8-Max wins seven.
Alibaba launched Qwen3.8-Max on Monday, calling it essentially the most succesful mannequin it has ever constructed. The weights land on Hugging Face and ModelScope subsequent week—the primary time the corporate has given away a mannequin at Max scale.
The specs are massive: 2.4 trillion parameters whole, with 95 billion switched on at any second. Parameters are the variety of dials a mannequin is ready to deal with.
That is extraordinarily vital for effectivity because it means it will not require too a lot sources to run. Consider it as an enormous library the place solely the related shelf lights up for every query. Small companies and labs with ok {hardware} are actually in a position to run a state-of-the-art mannequin with out spending the identical as an enormous datacenter.
Alibaba’s launch put up skips the standard benchmark-chasing narrative and leans on endurance as an alternative. The mannequin spent 16 days constructing a coding instrument by itself—265 commits, 127 pull requests, 151 points, no human touching the keyboard. It spent 5 days reproducing a analysis paper it had by no means seen code for, then beat the paper’s personal outcomes by 2.7 factors. In a 24-hour machine studying contest, it completed forward of 458 of 526 human groups.
Constructed to run inside a rival’s instruments
Qwen3.8-Max ships with directions for Claude Code and Codex, the coding instruments made by Anthropic and OpenAI. Alibaba’s API speaks each firms’ protocols. Most of its coding benchmarks had been run inside Claude Code.
And people benchmarks do not flatter it. Throughout 31 textual content exams, Anthropic’s Fable 5 takes 15 first place spots, OpenAI’s GPT-5.6 Sol takes 9, Qwen takes seven. On the 12 coding exams, Qwen wins precisely one. Nonetheless, by way of intelligence prices, this mannequin is extraordinarily low cost and environment friendly, which signifies that even when it requires extra iterations or reasoning, the price of getting the job accomplished might be a lot decrease, almost 30% of what Claude Fable 5 costs.

That stated, flip to multimodal work—paperwork, video, spatial reasoning—and the rankings invert. Qwen leads most of that desk.
There may be additionally a reversal by way of enterprise technique. In April, Alibaba killed the free tier of Qwen Code, with the crew drifting towards closed, paid fashions after management departures. Our evaluate of Qwen 3.7 Max famous the Plus model could be open whereas Max stayed locked behind the API.
That door is now open, and the timing is not an accident. Chinese language open-weight fashions went from underneath 2% of tokens on OpenRouter in late 2024 to roughly 61% by mid-2026. Qwen handed Meta’s Llama as essentially the most self-hosted mannequin on the earth.
In the meantime Washington restricted Fable 5 and Mythos 5 underneath export controls in June, and Beijing is reportedly weighing limits of its personal on Chinese language fashions going abroad.
So Alibaba is dropping on paper and successful on distribution. If you happen to can obtain one thing that comes shut totally free, second place is a high-quality place to be.
Each day Debrief E-newsletter
Begin each day with the highest information tales proper now, plus authentic options, a podcast, movies and extra.
