ChipsThe story, in brief

Nvidia Vera Rubin NVL72 Boosts Agentic AI Throughput by 30x

30x. Nvidia's Vera Rubin NVL72 claims a massive agentic AI throughput jump—but the fine print on power, cooling, and cost matters more than the headline.

Illustration of independent geometric mechanisms passing paper tasks along branching amber tracks.
AI agents and the coordination of work.AI illustration by KeyNews
The KeyNews take

Why it matters

Nvidia is optimizing its rack-scale silicon for agentic workloads, signaling a shift in how the compute buildout is being architected. Practitioners need to understand the real cost-per-inference and operational overhead before treating this as a drop-in win.

The key facts

9 to know
  1. Nvidia Vera Rubin NVL72 claims 30x agentic AI throughput improvement

  2. Throughput gain comes with power, cooling, and cost tradeoffs

  3. Agentic AI workloads reshaping hardware optimization priorities

  4. Published September 2026

  5. Nvidia Vera Rubin NVL72 announced

  6. Up to 30x agentic AI throughput improvement claimed

  7. Agentic AI workload optimization focus

  8. Power, cooling, and cost considerations flagged as critical trade-offs

  9. Enterprise evaluation needed for workload-specific ROI

Go to the source

TechRepublictechrepublic.com

Publisher excerpt: Nvidia says Vera Rubin NVL72 delivers up to 30x more agentic AI throughput, but enterprises should weigh workload needs, power, cooling, and cost.
Read original report
Back to today's editionMore chips news

The wider picture

View all
Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
AI illustration by KeyNews
Chips01

Why confidential computing is essential for enterprise AI

As enterprises push AI into sensitive domains (healthcare, finance, government), protecting data during processing — not just at rest or in transit — is shifting from a nice-to-have security feature to a prerequisite for deployment. HPE and NVIDIA are positioning confidential computing as foundational infrastructure for sovereign AI.

CIO
Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
AI illustration by KeyNews
Chips02

Beyond the limits of air: Why liquid cooling is becoming a strategic imperative for AI

Liquid cooling is shifting from niche to strategic necessity as GPU power density explodes. For practitioners building or deploying AI at scale, this isn't optional—it's a 18–24 month ROI play that unlocks denser racks, lower OpEx, and competitive advantage. The buildout architecture is changing.

CIO
Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
AI illustration by KeyNews
Chips03

Singapore’s Nexstrom wants to bring 2D semiconductors to chip fabs

2D materials promise better performance-per-watt for next-gen AI accelerators. Nexstrom's equipment funding signals the transition from lab to foundry—a potential shift in the compute buildout.

TechCrunch Startups