The Hallway Track
Product Launches

NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents

NVIDIA Developer Blog · Aug 11, 2026 · Product Launches

NVIDIA launches Nemotron 3.5 Lightning, a 30B MoE model optimized for agentic execution layers

“Long-running AI agents spend most of their time on high-volume execution: tool calls, result validation, and subagent delegation.”

NVIDIA released Nemotron 3.5 Lightning, an open 30B mixture-of-experts model with only 3B active parameters, purpose-built for the high-volume execution layer of long-running AI agents. The model targets tool calls, result validation, and subagent delegation — tasks where using frontier reasoning models is cost-prohibitive. This signals a maturing agentic AI stack where specialized, efficient models handle routine execution while larger models handle planning.

NVIDIA agentic-AI MoE inference-efficiency open-model

Watch / read the original source →