Why You Cant Tell When ChatGPT Is Wrong @WeightsBiases
Why You Cant Tell When ChatGPT Is Wrong  @WeightsBiases
Uploaded June 2026 | Updated September 2026, 2 weeks ago
What happens when you optimize your AI agent for customer satisfaction?

Say a shipping company deploys an LLM trained to get thumbs up. Someone calls asking where their lost package is. The system can admit it's lost or say it's coming tomorrow.

Saying the latter would make the customer happy and the agent would earn a thumbs up for lying.

Dan Klein on Gradient Dissent: that's not a bug, but a reward function working exactly as intended.
Why You Cant Tell When ChatGPT Is WrongW&B Mobile App for iOS is live!Your AI Agent is gaslighting you. Here are the receiptsThis Is Where AI Isn’t Trusted YetWhat a $42B Software Co. Really Spends on AI ToolsProtect your AI applications from risk and uncertainty with W&B Weave GuardrailsWhy Mathematicians Can’t Look Away From Axiom’s AILambda Labs David Hall on partnering with Weights & BiasesEvaluating AI applications using W&B WeaveCuring Every Disease With Al by 2050 | Sam Rodriques, Edison ScientificThe $2B Company Cutting AI Costs By 60% | Tuhin SrivastavaShe Raised $64M to Build an AI Math Prodigy | Carina Hong, CEO of Axiom
Weights & Biases |

Why You Can't Tell When ChatGPT Is Wrong

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER