FrontierThe story, in brief

DeepSeek's new models are so efficient they'll run on a toaster ... by which we mean Huawei's NPUs

DeepSeek V4 cuts inference costs to a fraction of R1 — now optimized for Huawei NPUs.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

DeepSeek's new efficiency gains signal a shift in the cost-performance equation for model deployment, with implications for who can afford to run frontier AI at scale.

The key facts

4 to know
  1. DeepSeek V4 now in preview

  2. Inference costs reduced significantly vs. R1

  3. Optimized for Huawei NPU hardware

  4. Efficiency improvements enable deployment on resource-constrained hardware

Go to the source

The Register AI/MLtheregister.com

Publisher excerpt: Now available in preview, DeepSeek V4 cuts inference costs to a fraction of R1
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier