Proving a Prompt Fix Works in Production with Phoenixs PXI @arizeai
Proving a Prompt Fix Works in Production with Phoenixs PXI  @arizeai
Uploaded August 2026 | Updated September 2026, 2 weeks ago
The experiment said the new prompt fixes the demo agent's SQL. But experiments aren't production! Watch the real spans come in and let Phoenix's evaluator be the judge.

Watch this video to learn:
• Why an experiment result isn't the finish line — production is
• The full loop: trace → dataset → evaluator → experiment → verify in production
• How PXI helped at every step: building the evaluators, the datasets, and the experiments

#ArizeAI #LLMEvaluation #AgentObservability #Phoenix

🎵 Music by Phoenix engineer Tony Powell — check out more of his music at soundcloud.com/cephalization/sets/some-sounds

🔗 Try Phoenix: arize.com/phoenix/?utm_source=youtube&utm_medium=video&utm_campaign=devrel&utm_content=rnabors
🔔 Subscribe for weekly content on LLMs, agents, and evaluation: youtube.com/@arizeai?sub_confirmation=1
Proving a Prompt Fix Works in Production with Phoenixs PXIHow PromptQL Built a Self-Updating Company Brain for AI Agents | Arize Observe 2026Harnessing User Feedback at ChatGPT Scale | OpenAI | Arize Observe 2026Building, Deploying, and Optimizing AI Agents with Microsoft Foundry | Arize Observe 2026Session Evaluation On An AI Tutor ChatbotMeet Quiet-STaR and Minimo: Understanding Self Discovered Reasoning EnvironmentsRise of the Agent Engineer: Booking.coms Chana RossHow LG Uplus Built an AI Contact Center Serving 30 Million Customers | Arize Observe 202The Math Humans Physically Cant DoKeynote | The Future of AI Agents | Arize Observe 2026AI Agent Got the Right Answer the Wrong Way | Rise of the AI Engineer | Michael Grinich, WorkOSUltimate OpenTelemetry Guide for Tracing AI Applications
Arize AI |

Proving a Prompt Fix Works in Production with Phoenix's PXI

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER