How to build a custom vision agent @googlecloudtech
How to build a custom vision agent  @googlecloudtech
Uploaded June 2026 | Updated September 2026, 2 weeks ago
Join Gemini Enterprise Agent Ready (GEAR) for the latest agent resources → https://goo.gle/4xpCT5W
John's GitHub repo → https://goo.gle/4uuvsYi

Ever wanted to turn your webcam into a generative media engine? Google Developer Expert, John Capobianco shares how he created a custom vision agent using the Google Gemini ecosystem and MCP, capturing live photos with Nano Banana Pro and animating them using Veo 3. It even processes natural language prompts and real-time ASL conversations.

Speakers: John Capobianco
Products Mentioned: Gemini, Model Context Protocol, Nano Banana, Veo 3
How to build a custom vision agentAI-powered observability & remediation with Google Gemini CLI and DynatraceJoin AI Agent Clinic on April 15thScale JAX models to multi-GPU systemsGenerative UI for any agent, anywhere: A2UI, AG-UI, MCP Apps, and moreBuild long-running agents with Google’s Agentic Stack | The Agent Factory5 agent patterns to masterHow Google Developer Experts vibecoded an AI racing coach with GeminiAlloyDB trial clusters: your playground for PostgreSQL innovationBuilding distributed multi-agent systemsBuild your first Android app in AI Studio in 5 minutesData agent kit: Your coding agent can now query your data
Google Cloud Tech |

How to build a custom vision agent

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER