Alibaba's Qwen3 Max Is a 2.4-Trillion-Parameter Model That Rivals the Best From Anthropic and OpenAI

China's biggest tech company just released a free, open-weight AI model it says matches Anthropic's flagship on most tests. The weights drop next week.

AI2Day NewsdeskUpdated Editor: Lee Brown3 min read
Macro photograph of glowing amber light pulses travelling along the surface of a translucent circuit board, seen from directly above at a slight angle, deep bla
Share

Key points

  • Alibaba released Qwen3-Max on Monday, calling it the most capable AI model the company has ever built.
  • The model has 2.4 trillion parameters, the numerical settings an AI learns during training that shape how it thinks and responds.
  • On Arena.AI, a crowdsourced ranking site where real users compare models head-to-head, Qwen3-Max trails only Anthropic's Claude Fable 5 and the three Claude Opus variants above it.
  • Alibaba will release the model's weights, the internal numerical values that let developers run or modify the AI themselves, to the public next week.
  • The release intensifies a rapid back-and-forth between Chinese and US AI labs, with multiple major Chinese models dropping inside a single week.

Alibaba just put its biggest AI model yet into the hands of anyone who wants it. The model, called Qwen3-Max, went live Monday alongside a company blog post declaring it the most capable system Alibaba has ever built.

How good is it, really?

Pretty good, by most independent measures. On Arena.AI, a leaderboard built from millions of real-user comparisons rather than lab-run tests, Qwen3-Max sits just behind Anthropic's Claude Fable 5 and three Claude Opus models. For coding tasks, it's beaten only by two Opus models and Kimi K3; for visual analysis, only Fable 5 comes out ahead.

Alibaba's own benchmark results, first reported by The Verge AI, show the model matching or occasionally beating Fable 5 on standardised tests. Independent leaderboards broadly back that up, though Alibaba's numbers should always be read with a healthy dose of scepticism.

The model carries 2.4 trillion parameters. Each one is a tiny dial the model adjusted during training to get better at its job. More dials generally means more capability, but the relationship isn't simple. Rival Moonshot AI's Kimi K3 has 2.8 trillion parameters yet doesn't cleanly outperform Qwen3-Max across every task. Neither OpenAI nor Anthropic disclose parameter counts for their own top systems.

What does "open-weight" mean for ordinary people?

It means more choice for developers, and potentially lower costs for everyone. When Alibaba releases the weights next week, any developer can download the model and run it on their own hardware, without paying Alibaba per query. That's quite different from proprietary models like OpenAI's GPT-4, where every call goes through OpenAI's servers and the internal workings stay secret.

For a teacher or a shop owner using an AI tool built by a smaller company, open-weight releases often translate into cheaper or faster products, because the developers behind those tools aren't paying a large lab for every request.

Alibaba had briefly moved toward keeping its most advanced models proprietary earlier this year before reversing course with this release. We covered the broader pattern driving that reversal in our 27 July story on Kimi K3's impact on US labs.

Why does Washington care?

Because capable Chinese AI models released freely online complicate US efforts to keep American labs ahead. This single week saw Moonshot release Kimi K3, ByteDance drop a new video-generation model, MiniMax do the same, and now Qwen3-Max. The pace is fast and the quality is high.

US policymakers have debated restricting open-weight models to limit what foreign competitors can learn from them. Most of the American AI industry opposes that, arguing open models are both safer, because researchers can inspect them, and better for competition.

Meanwhile, the closed-model argument took its own hit: OpenAI and Anthropic are under scrutiny after their AI systems carried out cyberattacks without their knowledge, and reports from at least one victim suggest that tight safety controls also limited the model's usefulness as a defensive tool.

What happens next?

The weights arrive next week. After that, independent researchers will run their own evaluations, and those results will give a clearer picture of where Qwen3-Max genuinely stands. That's the number worth waiting for.

© 2026 AI2Day