TLDR

  • Alibaba’s Qwen team previewed Qwen3.8-Max-Preview on July 19 with 2.4 trillion parameters, claiming it ranks “second only” to Anthropic’s Claude Fable 5
  • The flagship is the first Qwen model above 1 trillion parameters and is fully multimodal, processing text, images, video, and documents
  • Preview access runs through Alibaba’s Token Plan, Qoder, and QoderWork at 10% of standard pricing — open weights promised “soon” with no license yet
  • Drop landed during Shanghai’s World AI Conference (WAIC), two days after Moonshot AI released Kimi K3, a 2.8 trillion-parameter open-weight model
  • Active parameters per token are undisclosed, so serving cost for a 2.4T open-weight release remains unknown — community mood is cautiously optimistic

Alibaba Enters The Multi-Trillion Parameter Race

image of Alibaba’s Qwen3.8-Max Hits 2.4 Trillion Parameters, Claims Second Place Behind Anthropic’s Fable 5 - HelloExpress - 2

Alibaba dropped Qwen3.8-Max-Preview on July 19, 2026 during the World AI Conference (WAIC) in Shanghai, and the Chinese tech giant did not bother with understatement. The official Qwen X account called the model “one of the most powerful available today, comparable to leading frontier AI models, second only to Fable 5,” referring to Anthropic’s most capable widely released model. With 2.4 trillion parameters, Qwen3.8-Max-Preview is the first Qwen model to cross the trillion-parameter line, and Alibaba is using that milestone to argue Chinese AI labs have entered the “multi-trillion parameter age.”

image of Alibaba’s Qwen3.8-Max Hits 2.4 Trillion Parameters, Claims Second Place Behind Anthropic’s Fable 5 - HelloExpress - 3

The release lands two days after Beijing-based Moonshot AI unveiled Kimi K3, a 2.8 trillion-parameter open-weight model that holds the title of the world’s largest open-source AI system. The back-to-back timing reads as a coordinated signal from China’s AI sector: the frontier is no longer a Western-only conversation, and the parameter arms race has shifted into a new gear that US labs will struggle to match on raw scale alone.

What The Preview Actually Gives Developers

Qwen developer Shuai Bai confirmed on X that Qwen3.8 is the team’s first multimodal model above one trillion parameters, handling text, images, video, and documents in a single system. According to the team, the new flagship should beat Qwen3.7-Max on coding, full-stack development, data analysis, and office workflows — a productivity layer Alibaba is pitching, not a research demo.

image of Alibaba’s Qwen3.8-Max Hits 2.4 Trillion Parameters, Claims Second Place Behind Anthropic’s Fable 5 - HelloExpress - 3

Access is live now through three channels. Alibaba’s Token Plan subscription offers the preview at 10% of standard pricing, and the same build is wired into Qoder and QoderWork, Alibaba’s agentic coding platforms. Open weights are promised “soon,” but no release date or license have been published. Until the Hugging Face repository lands with a real license, production deployments should stay on the official API.

Why The 2.4 Trillion Number Is Misleading

Total parameter count is not the same as usable compute, and Qwen’s own history proves the point. Earlier Qwen3 models used sparse mixture-of-experts designs where active parameters per token were a tiny fraction of total — Qwen3-235B-A22B activated just 22 billion of its 235 billion per token. Qwen3.8-Max-Preview is widely assumed to follow the same MoE pattern, but the active parameter count has not been disclosed.

That matters because serving cost is driven by active parameters, not the headline number. Startup Fortune calculated that a 2.4T model at 4-bit precision needs roughly 1.2 terabytes for weights — a single Nvidia H200 carries 141GB. Without a smaller activated-parameter variant, a quantized checkpoint, or a distilled sibling, the open-weight promise may only matter to well-funded hyperscalers and the largest Malaysian cloud buyers. For everyone else, the API is the only realistic door in.

Our Take

The interesting part of Qwen3.8-Max-Preview is not whether Alibaba’s “second only to Fable 5” claim holds up — that benchmark is Alibaba’s own, and independent evaluation will take weeks. What matters is the structural shift: a credible Chinese flagship has landed in the same parameter range as US frontier models, with a multimodal build and a discount-priced API that undercuts Western rivals per token. For Malaysian developers, that is meaningful — Alibaba Cloud’s Kuala Lumpur region serves regional customers, and a cheap, capable Chinese model is a real alternative to OpenAI or Anthropic for production workloads that do not require the absolute frontier.

The honest read, though, is that this is a teaser, not a release. The benchmark table is Alibaba’s, the license is missing, the weights are promised, and the active-parameter count is undisclosed. Community sentiment on Hacker News and Reddit r/LocalLLaMA split cleanly: pro-open-weight competition, but no one is migrating production traffic on an X post. The right move is to test Qwen3.8-Max-Preview through the official console, wait for the open-weight drop and independent evaluation, and only then decide if it earns a slot next to Claude, GPT, or Gemini. The frontier just got more crowded — good news for buyers, even if the marketing is louder than the math.

Source

You may also like

Leave a reply

Your email address will not be published. Required fields are marked *