The Hallway Track

voice-agents

14 tracked signals on voice-agents.

Voice Agent observability with LangSmith 🌟

Google Developers (Google I/O) · Aug 05, 2026

LangSmith now provides full observability for voice agents built on Gemini Live

“To take that agent to production safely, you need visibility into what your agent is doing.”
Evaluate your Amazon Nova Sonic voice agent at scale, no microphone required

AWS Machine Learning Blog · Jun 08, 2026

AWS released the open source Nova Sonic Test Harness to automatically evaluate voice agents at scale without a microphone.

“It runs complete multi-turn conversations with Amazon Nova Sonic automatically, evaluates them using LLM-as-judge techniques, and can even detect cases where the model's audio output doesn't match its text output (audio hallucinations).”
Build realtime multimodal agents with LiveKit and Azure | ODSP937

Microsoft Developer (Build) · Jun 03, 2026

Real-time multimodal voice agents are bottlenecked by media infrastructure, not models, which LiveKit handles via WebRTC on Azure.

“You see models they are the easy part. Now everything around it is the hard part.”
Voice Agent observability with LangSmith

Google Developers (Google I/O) · Jul 31, 2026

Google ADK voice agents using Gemini Live can be traced with LangSmith for production observability

“To take that agent to production safely, you need visibility into what your agent is doing and the ability to test its behavior in a number of different scenarios that it might encounter with real end users.”
Build a healthcare appointment agent with Amazon Nova 2 Sonic

AWS Machine Learning Blog · Jun 24, 2026

AWS shows how to build a healthcare voice appointment agent using Amazon Nova 2 Sonic and Bedrock AgentCore.

“Instead of chaining separate transcription, reasoning, and synthesis services, Nova 2 Sonic processes speech natively in a single model—so vocal context like tone and pace isn’t lost to transcription.”
Turn Any LangGraph Agent Into a Voice Agent in Minutes

LangChain · Jun 23, 2026

LangChain shows how to convert an existing LangGraph agent into a voice agent using the Pipecat framework.

“Pipecat is going to handle all of the glue required to take the input audio, convert it to text, run it through our LangGraph LLM layer, then convert that back to speech in order to send it back to the end user.”
Voice for AI Agents and Applications

DeepLearningAI · Jun 17, 2026

DeepLearning.AI launched a course on building fast, reliable voice agents in partnership with Vocal Bridge.

“Voice is an under-exploited frontier.”