DeepSeek's new models are so efficient they'll run on a toaster ... by which we mean Huawei's NPUs
DeepSeek V4 cuts inference costs to a fraction of R1 — now optimized for Huawei NPUs.

Why it matters
DeepSeek's new efficiency gains signal a shift in the cost-performance equation for model deployment, with implications for who can afford to run frontier AI at scale.
The key facts
4 to knowDeepSeek V4 now in preview
Inference costs reduced significantly vs. R1
Optimized for Huawei NPU hardware
Efficiency improvements enable deployment on resource-constrained hardware
Go to the source
The Register AI/MLtheregister.com
Publisher excerpt: Now available in preview, DeepSeek V4 cuts inference costs to a fraction of R1