End-to-End Testing With AI Agents: What I Built, Broke, and Learned @DevOpsToolkit
End-to-End Testing With AI Agents: What I Built, Broke, and Learned  @DevOpsToolkit
Uploaded August 2026 | Updated September 2026, 4 days ago
AI tools can now write your tests, maintain them, and even diagnose failures — but handing all of that to an agent introduces a risk that's easy to miss. When AI writes bad code, your tests catch it. When AI writes bad tests, nothing does. They just pass, and you become confident about something that isn't true.

This video breaks down what's actually happening when we talk about "testing" — drawing a sharp line between *checking* (confirming what you already believe) and *testing* (going and finding out what you got wrong). Through a real demonstration using Momentic, it explores how AI agents can generate end-to-end test suites from a single instruction, self-repair failing tests, and even explore a product autonomously to surface bugs nobody specified in advance. The uncomfortable conclusion: the judgment, the suspicion, the poking-at-things-to-see-what-breaks — the part everyone assumed would always need a human — is no longer safely in human hands. What remains is the one thing no agent can supply on its own: deciding what the software is actually supposed to do, and making sure that intent gets into the system before anything else runs.

#AITesting #TestAutomation #SoftwareQuality

Consider joining the channel: youtube.com/c/devopstoolkit/join

▬▬▬▬▬▬ 🔗 Additional Info 🔗 ▬▬▬▬▬▬
➡ Transcript and commands: https://devopstoolkit.live/development/ai-testing-is-lying-to-you-and-you-cant-tell
🔗 Momentic: https://momentic.plug.dev/aLtwFNd

▬▬▬▬▬▬ 💰 Sponsorships 💰 ▬▬▬▬▬▬
If you are interested in sponsoring this channel, please visit https://devopstoolkit.live/sponsor for more information. Alternatively, feel free to contact me over Twitter or LinkedIn (see below).

▬▬▬▬▬▬ 👋 Contact me 👋 ▬▬▬▬▬▬
➡ BlueSky: https://vfarcic.bsky.social
➡ LinkedIn: linkedin.com/in/viktorfarcic

▬▬▬▬▬▬ 🚀 Other Channels 🚀 ▬▬▬▬▬▬
🎤 Podcast: devopsparadox.com
💬 Live streams: youtube.com/c/DevOpsParadox

▬▬▬▬▬▬ ⏱ Timecodes ⏱ ▬▬▬▬▬▬
00:00 Testing with AI
01:51 AI Writes And Repairs Tests
05:45 Why Failing Tests Go Green
08:00 Testing Versus Checking
11:25 Why QA Engineers Disappeared
15:27 What's Left For Humans
End-to-End Testing With AI Agents: What I Built, Broke, and LearnedHow I Went From Writing Code to Managing AI Agents Full-TimeEp33 - Ask Me Anything About Anything with Scott RosenbergEp10 - Ask Me Anything About DevOps, Cloud, Kubernetes, Platform Engineering,... w/Scott RosenbergMCP Servers Explained: Why Most Are Useless (And How to Fix It)Ep22 - Ask Me Anything About Anything with Scott RosenbergDevOps Q&A: Crossplane, AI Agents, GitOps Terraform, and Career TipsAI doesnt have to do reviews perfectlyDevOps Q&A: MCP Servers, Kubernetes Observability, and Progressive DeliveryMe vs. AI - Live Match Against GitHub Copilot AgentEp31 - Ask Me Anything About Anything with Scott RosenbergAI Meets Kubernetes: Simplifying Developer and Ops Collaboration
DevOps & AI Toolkit |

End-to-End Testing With AI Agents: What I Built, Broke, and Learned

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER