ChipsThe story, in brief

NVIDIA CEO Jensen Huang at Dell Technologies World: “Demand Is Going Parabolic, Utterly Parabolic”

One-tenth the cost per token. NVIDIA's Vera Rubin just reset the economics of agentic AI inference.

Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
The infrastructure powering AI.AI illustration by KeyNews
The KeyNews take

Why it matters

NVIDIA is shipping hardware (Vera Rubin NVL72) that materially reduces inference costs and speeds up agent workloads by 50%, with 5,000 enterprises already deployed. This shifts the compute economics for enterprise AI deployment and signals where the infrastructure battle is heading.

The key facts

6 to know
  1. NVIDIA Vera Rubin NVL72: agentic AI inference at 1/10th cost per token

  2. Agent sandboxes run 50% faster on Vera vs traditional CPUs

  3. Enterprise data queries up to 3x faster with Vera CPU

  4. 5,000 enterprises deployed: Lilly, Samsung, Honeywell, others

  5. Dell AI Factories integration narrative

  6. Jensen Huang quote on parabolic demand trajectory

Go to the source

NVIDIA Blogblogs.nvidia.com

Publisher excerpt: Agentic AI inference at one-tenth the cost per token with NVIDIA Vera Rubin NVL72. Agent sandboxes run 50% faster on NVIDIA Vera than traditional CPUs — while enterprise data queries are up to 3x faster with the Vera CPU. And 5,000 enterprises like Lilly, Samsung, and Honeywell are running AI…
Read original report
Back to today's editionMore chips news

Keep reading

Related stories

More from Chips