FrontierThe story, in brief

NVIDIA AI Releases Nemotron 3.5 Lightning: A 30B Open MoE with 3B Active Parameters, and NeMo Switchyard Model Router

NVIDIA open-sources a 30B MoE with just 3B active parameters—and ships a router that picks the cheapest capable model for each step.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

A major open-weight release targeting agent execution efficiency. Nemotron 3.5 Lightning and Switchyard together signal NVIDIA's play to own the inference layer for agentic workloads—competing directly with smaller, cost-optimized models from rivals.

The key facts

6 to know
  1. Nemotron 3.5 Lightning: 30B parameters, 3B active (MoE architecture)

  2. Open-weight release (no fence)

  3. NeMo Switchyard: model router for agent execution layer

  4. Routing strategy: selects cheapest capable model per step

  5. Agent execution framed as primary use case

  6. Published August 11, 2026

Go to the source

MarkTechPostmarktechpost.com

Publisher excerpt: NVIDIA's open 30B MoE targets the agent execution layer, with Switchyard routing each step to the cheapest capable model.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier