NVIDIA CEO Jensen Huang at Dell Technologies World: “Demand Is Going Parabolic, Utterly Parabolic”
One-tenth the cost per token. NVIDIA's Vera Rubin just reset the economics of agentic AI inference.

Why it matters
NVIDIA is shipping hardware (Vera Rubin NVL72) that materially reduces inference costs and speeds up agent workloads by 50%, with 5,000 enterprises already deployed. This shifts the compute economics for enterprise AI deployment and signals where the infrastructure battle is heading.
The key facts
6 to knowNVIDIA Vera Rubin NVL72: agentic AI inference at 1/10th cost per token
Agent sandboxes run 50% faster on Vera vs traditional CPUs
Enterprise data queries up to 3x faster with Vera CPU
5,000 enterprises deployed: Lilly, Samsung, Honeywell, others
Dell AI Factories integration narrative
Jensen Huang quote on parabolic demand trajectory
Go to the source
NVIDIA Blogblogs.nvidia.com
Publisher excerpt: Agentic AI inference at one-tenth the cost per token with NVIDIA Vera Rubin NVL72. Agent sandboxes run 50% faster on NVIDIA Vera than traditional CPUs — while enterprise data queries are up to 3x faster with the Vera CPU. And 5,000 enterprises like Lilly, Samsung, and Honeywell are running AI…