The Hallway Track

NVIDIA GTC Spring 2026

inference robotics hardware

Dates
2026-03-16 → 2026-03-20
Location
San Jose, CA
Ecosystem
chip infra
Importance
9/10

Related coverage & signals

Inside NVIDIA Rubin GPU Architecture: Powering the Era of Agentic AI

NVIDIA Developer Blog · Jul 21, 2026

NVIDIA Rubin GPU architecture is purpose-built for agentic AI workloads at scale

“What began as discrete AI model training and human-facing chat interfaces has evolved into always-on AI factories dedicated to producing intelligence at scale.”
Cerebras Supernova 2026

Cerebras · Aug 21, 2026

Cerebras claims it has eliminated the speed-throughput tradeoff in AI inference

“In AI, speed is productivity.”
You're Underestimating America

No Priors · Aug 02, 2026

US AI supply chain faces extreme concentration risk beyond chips, especially in robotics components dominated by China

“Robotics is an incredibly promising industry and the supply chain is right now completely dominated by China.”
Scale AI with Google's TPU software stack

Google Developers (Google I/O) · May 21, 2026

Google reveals TPU V8 splits into training (8T) and inference (8I) specialized variants

“a lot of the intelligence is actually coming from inference”
Cerebras Explains | What is Fast AI Inference?

Cerebras · Sep 25, 2026

Cerebras runs models at over 1000 tokens per second for near-instantaneous inference

“We run models at over 1000 tokens per second, so thinking seems instantaneous.”
Cerebras Explains | What Is LLM Inference?

Cerebras · Sep 30, 2026

LLM inference is bottlenecked by memory bandwidth, not compute speed

“LLM inference is actually limited by memory bandwidth, not computational speed.”
[AINews] AMD buys Taalas

Lisa Su · Latent Space Blog · Aug 07, 2026

AMD acquires Taalas, signaling Lisa Su's conviction in custom ASICs for AI inference

[AINews] Memory prices up 500% in 12 months

Latent Space Blog · Aug 19, 2026

Memory prices up 500% in 12 months as hyperscalers lock in all 2027 DRAM production capacity

“Some are calling it the RAMpocalypse; I prefer "RAMageddon."”
Introducing Core AI

Apple Developer (WWDC) · Aug 17, 2026

Apple launches Core AI framework for on-device model inference across iOS and macOS.

Introducing Gemini Robotics 2

Google Developers (Google I/O) · Jul 31, 2026

Google DeepMind launches Gemini Robotics 2 with whole-body intelligence, dexterity, and multi-robot collaboration.

“What we're building here is the intelligence layer to power any robot to do a broad range of useful tasks.”
OpenAI is being sued for stealing, again…

Sam Altman · Fireship · Jul 17, 2026

Apple sued OpenAI for trade secret theft tied to its $6.5B hardware push

“OpenAI believes the product's defining feature will be its personality and ability to connect on a human-like level with users.”
Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era

Jack Clark · Import AI · Jun 29, 2026

NVIDIA's ENPIRE framework lets coding agents run a closed-loop self-improvement process for real-world robots, hitting 99% on dexterous tasks.

“Frontier coding agents can autonomously develop a policy to achieve a 99% success rate on challenging, dexterous manipulation tasks in the real world, such as PushT, organizing pins into a pin box, and using a cutter to cut a zip tie,”
From VLM/VLA's to Embodied Agents — Armen Aghajanyan, Perceptron AI

AI Engineer · Sep 23, 2026

Perceptron AI wants to unify VLMs, VLAs, and world models into 'embodied foundational models' for real-time physical-world interaction.

“we want to move away from the distinction between VLM, VLA, world models, whatever you want to call it, to what we call embodied fundamental models”