Building Voice Agents with Gemini 3 @GoogleDevelopers
Building Voice Agents with Gemini 3  @GoogleDevelopers
Uploaded March 2026 | Updated September 2026, 1 week ago
*Build real-time conversational agents with Gemini 3*
Thor from Google DeepMind walks through the Gemini Live API, showing how to build natural, human-like voice interactions powered by Gemini’s native audio model: speech-to-speech, no text in the middle, with emotional nuance, multilingual support, and real-time tool use.

*What’s covered:*
Testing in Google AI Studio, streaming audio and video frames, configuring voices and system instructions, WebSocket integration with the GenAI SDK, session management, interruption handling, and deploying with partner frameworks like LiveKit, Daily, and Stream.

Get started:
Try Gemini Live in Google AI Studio and grab your API key to start building.

Resources:
✅Live API documentation → https://goo.gle/4rObPZV
✅GitHub examples → https://goo.gle/4c4sZhi
✅Blog post → https://goo.gle/4m1KoLa

What are you building with Gemini Live? Drop it in the comments.


Subscribe to Google for Developers → https://goo.gle/developers

Speaker: Thor Schaeff
Products Mentioned: Gemini, Google AI, Google AI Studio
Building Voice Agents with Gemini 3Gemma Playground: AI Edge GalleryHow do AI video generation models work?Add databases to your app with AI Studio | Vibe Coding GuideADK for Java 1.0 is now available!Sameer Samat on Android 17 and the Future of Intelligent ComputingHow to build a full-stack app with Supabase and Stripe on Google AI StudioBuild a live translation broadcast app with the Gemini Live API and LiveKitCan you spot which committed file explains this mismatch?Introducing Keras Recommenders: state-of-the-art recommendation techniques at your fingertipsSee how Gemma  can explore, plan, and scale!Create advanced data driven Gemini API apps
Google for Developers |

Building Voice Agents with Gemini 3

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER