Today - Monday Aug 3

Alibaba Ships Its Biggest AI Model

3AUG
#2 GLOBALALIBABACLAUDE

Alibaba unveiled Qwen3.8-Max, its largest model ever, on Monday. The 2.4 trillion parameter model ranks second globally on image benchmarks. It still trails Claude on text, and full release lands next week.

A mixture of experts design keeps costs down. Only ninety five billion of the total parameters activate per request. That's how Alibaba keeps inference cheap at frontier scale.

Reuters frames this as a fierce race among Chinese firms building cheaper models. Its parameter count sits close to Moonshot's Kimi K3, which has two point eight trillion.

Alibaba hasn't published a full benchmark table yet. So today's numbers are still just the company's own claims. Independent testing will decide if that vision ranking actually holds.

full brief & sources

Why this matters

  • Chinese labs keep closing the gap with US frontier models, fast.
  • Parameter count and open weights are becoming Alibaba's key recruiting pitch to developers.
  • Cost matters as much as capability now that mixture-of-experts design cuts inference bills.

🔍 What happened

  • Alibaba unveiled Qwen3.8-Max on Monday, August 3.
  • The model has 2.4 trillion parameters, close to Moonshot's 2.8 trillion parameter Kimi K3.
  • Only 95 billion parameters activate per request under its mixture-of-experts design.
  • It ranks second globally on Arena.AI's image and video leaderboard, behind a Claude Fable 5 variant.
  • On text tasks it still trails Claude Fable 5 and three Anthropic Opus variants.
  • Full release through Alibaba Cloud's Model Studio is set for next week.

💬 Smart takes

  • Alibaba: the model completed a full software-engineering project in 16 days during internal testing.
  • Reuters: Chinese tech companies are "locked in a fierce and fast-moving battle" to build powerful models cheaply.
  • Skeptic: parameter count is a marketing number. Qwen3.8-Max still trails Claude on the benchmark that matters most, text reasoning.

🧭 Where this goes

  1. LikelyAlibaba leans on the vision leaderboard ranking as its main marketing hook once the model ships next week.
  2. LikelyUS labs keep their parameter counts secret, making direct comparisons harder to verify.
  3. Possibleindependent benchmarks show a smaller gap, or a bigger one, than Alibaba's own numbers suggest.
  4. Wild CardQwen3.8-Max's cost advantage pulls meaningful US enterprise workloads away from Anthropic and OpenAI within months.

🥄 The Spoon Take

Alibaba keeps playing the same card: bigger parameter count, lower price, open weights. It's working on developers even if text benchmarks still favor Claude. Watch the vision leaderboard ranking, not the headline parameter count. That's where Qwen3.8-Max actually earned second place.

🤔 Pushback

Alibaba hasn't published a benchmark table yet, so every number here is still Alibaba's own claim.