How Evals and Prompts Shape Agent Behavior — Preetika Bhateja & Daniel Bump, YouTube Ads @aiDotEngineer
How Evals and Prompts Shape Agent Behavior — Preetika Bhateja & Daniel Bump, YouTube Ads  @aiDotEngineer
Uploaded July 2026 | Updated September 2026, 3 weeks ago
Getting an AI agent to behave the way you want isn't just about writing better prompts. In real systems, behavior emerges from a loop: prompts, evals, iteration, and feedback. Small changes in any part of that loop can completely change outcomes.

The Google team shares lessons from building a seed-asset agent that turns messy advertising creatives — low-quality images, cluttered visuals, and heavy text overlays — into clean, reusable assets for downstream generative AI tools. They explain why prompting alone did not produce stable behavior, how evals became feedback signals rather than scorecards, how agent trace logs exposed why failures happened, and how they iterated without breaking problems they had already fixed.

Speakers:

Chris Souza — Google
Chris works on the Google team behind this seed-asset agent and its evaluation workflow.

Preetika Bhateja — Product Manager, Google/YouTube
Preetika works on ads, evaluations, agents, and LLM-as-judge systems.

Daniel Bump — Engineer, Google
Daniel focuses on image and video generation and computer vision.
X/Twitter: https://x.com/DanielJBump
LinkedIn: linkedin.com/in/danielbump
How Evals and Prompts Shape Agent Behavior — Preetika Bhateja & Daniel Bump, YouTube AdsAI is the World’s largest Relationship Therapist — Clay Cockrell & Tony Fabrikant, CoupleWork AIInside 847 Production Clinical AI Notes — Sebastian Fox, ComposoYour Finance Agents Bottleneck Is You — Ramana Siddanth Emani, Auditoria AIVoice agents with Realtime Video — Sidney Primas, LemonSliceCoding Agents Dont Scale Themselves. Neither Do Your Teams. — Patrick Debois, TesslScaling up Continual Learning — Ronak Malde, TrajectoryAgents Need Feature Flags - Sachin GuptaUnlock Agent Autonomy: The Runtime for AI-Native Systems — Tushar Jain, DockerEmulated: The Data for Fully Autonomous Software Engineers and Companies — Joseph WangProductionizing LLM Gateways: Architecture, Tradeoffs and Hard Lessons — Kanish Manuja, Twilio
AI Engineer |

How Evals and Prompts Shape Agent Behavior — Preetika Bhateja & Daniel Bump, YouTube Ads

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER