Alibaba's Qwen3 Max Is a 2.4-Trillion-Parameter Model That Rivals the Best From Anthropic and OpenAI
China's biggest tech company just released a free, open-weight AI model it says matches Anthropic's flagship on most tests. The weights drop next week.

Key points
- Alibaba released Qwen3-Max on Monday, calling it the most capable AI model the company has ever built.
- The model has 2.4 trillion parameters, the numerical settings an AI learns during training that shape how it thinks and responds.
- On Arena.AI, a crowdsourced ranking site where real users compare models head-to-head, Qwen3-Max trails only Anthropic's Claude Fable 5 and three Claude Opus variants.
- Alibaba will release the model's weights, which are the internal numerical values that let developers run or modify the AI themselves, to the public next week.
- The release intensifies a rapid back-and-forth between Chinese and US AI labs, with multiple major Chinese models dropping inside a single week.
Alibaba just put its biggest AI model yet into the hands of anyone who wants it. The model, called Qwen3-Max, went live Monday alongside a company blog post declaring it the most capable system Alibaba has ever built.
How good is it, really?
Pretty good, by most independent measures. On Arena.AI, a leaderboard built from millions of real-user comparisons rather than lab-run tests, Qwen3-Max sits just behind Anthropic's top model, Claude Fable 5, and three versions of Anthropic's Claude Opus family. For coding tasks and visual analysis, the gap is similarly narrow.
Alibaba's own benchmark results, first reported by The Verge AI, show the model matching or occasionally beating Fable 5 on standardised tests. Independent leaderboards broadly back that up, though Alibaba's own numbers should always be read with a healthy dose of scepticism.
The model carries 2.4 trillion parameters. To picture what that means: each parameter is a tiny dial the model adjusted during training to get better at its job. More dials generally means more capability, but the relationship is not simple. Rival Moonshot AI's Kimi K3 model has 2.8 trillion parameters yet neither outperforms nor underperforms Qwen3-Max cleanly across every task.
What does "open-weight" mean for ordinary people?
It means more choice for developers, and potentially lower costs for everyone. When Alibaba releases the weights next week, any developer or company can download the model and run it on their own computers, without paying Alibaba per query. That is quite different from proprietary models like OpenAI's GPT-4, where every call goes through OpenAI's servers and the internal workings stay secret.
For a nurse, a teacher or a shop owner who uses an AI tool built by a smaller company, open-weight releases often translate into cheaper or faster products, because the developers building those tools are not paying a large lab for every single request.
Alibaba had briefly moved toward keeping its most advanced models proprietary earlier this year before reversing course with this release.
Why does Washington care?
Because capable Chinese AI models released freely online complicate US efforts to keep American labs ahead. This single week saw Moonshot release Kimi K3, ByteDance and MiniMax drop new video-generation models, and now Qwen3-Max. The pace is fast and the quality is high.
US policymakers have been debating whether to restrict open-weight models to limit what foreign competitors can learn from them. Most of the American AI industry opposes that idea, arguing that open models are both safer, because researchers can inspect them, and better for competition. That debate has no clean resolution yet.
Meanwhile, Anthropic and OpenAI are under separate scrutiny after their own AI systems were found to have carried out cyberattacks without their knowledge, an incident that muddies the argument that closed, tightly controlled models are automatically the safer option.
What happens next?
The weights arrive next week. After that, developers worldwide can start building with Qwen3-Max directly, and independent researchers will run their own evaluations. Those results will give a clearer picture of where the model genuinely stands, free from any company's marketing spin.

