The Hallway Track

gpu

10 tracked signals on gpu.

Amazon SageMaker Inference: 2026 year-to-date launches in review

AWS Machine Learning Blog · Sep 18, 2026

SageMaker AI shipped 13 new inference capabilities in 2026 across managed endpoints and HyperPod.

“Generative AI inference is uniquely hard: models are tens to hundreds of gigabytes, latency requirements are measured in tokens per second, cold starts can span multiple minutes as containers and weights transfer, GPU capacity is constrained, and traditional monitoring tools expose none of the token-level signals that matter in production.”
Manage Kubernetes Node Fleets with NodeWright

NVIDIA Developer Blog · Sep 23, 2026

NVIDIA introduces NodeWright to manage Kubernetes node fleets for GPU workloads.

“Kubernetes manages what runs on your nodes. Managing the nodes themselves is the challenge”
CCCL Runtime: A Modern C++ Runtime for CUDA

NVIDIA Developer Blog · Jun 22, 2026

NVIDIA introduces CCCL Runtime, modern C++ abstractions for core CUDA programming concepts.

“NVIDIA CUDA Core Compute Libraries (CCCL) provides delightful and efficient abstractions for CUDA developers in C++ and Python.”