FrontierThe story, in brief

Alibaba's latest AI model ran autonomously for 35 hours to optimize code for its own custom chip

35 hours. That's how long Alibaba's Qwen3.7-Max ran autonomously optimizing code for its own chip—and it's now matching Claude Opus on benchmarks.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Alibaba's Qwen3.7-Max demonstrates a new frontier in autonomous agent capability and long-context reasoning, positioning Chinese AI labs as competitive with frontier Western models on reasoning tasks while achieving practical chip-design applications.

The key facts

7 to know
  1. Qwen3.7-Max matches Claude Opus 4.6 on benchmarks

  2. Model ran autonomously for 35 hours optimizing code for custom chip

  3. Beats DeepSeek V4 Pro and Kimi K2.6 on tested benchmarks

  4. Proprietary model built for long-running autonomous agent tasks

  5. Demo includes four-legged robot control capability

  6. Released by Alibaba's Qwen team

  7. Published May 23, 2026

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: Alibaba's Qwen team releases Qwen3.7-Max, a proprietary model built for long-running autonomous agent tasks. It matches Claude Opus 4.6 on benchmarks and beats Chinese rivals like DeepSeek V4 Pro and Kimi K2.6. The team also demos the model steering a four-legged robot.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier