Uploaded May 2025 | Updated September 2026, 3 weeks ago
Following the theme of AI research and safety, Aric Floyd talks about how some Large Language Models might follow the all too human trait of sandbagging - "lying" about their true capabilities.
AI Sandbagging Paper: apolloresearch.ai/research/scheming-reasoning-evaluations
Computerphile is supported by Jane Street. Learn more about them (and exciting career opportunities) at: jane-st.co/computerphile
This video was filmed and edited by Sean Riley.
Computerphile is a sister project to Brady Haran's Numberphile. More at bradyharanblog.com
Following the theme of AI research and safety, Aric Floyd talks about how some Large Language Models might follow the all too human trait of sandbagging - "lying" about their true capabilities.
AI Sandbagging Paper: apolloresearch.ai/research/scheming-reasoning-evaluations
Computerphile is supported by Jane Street. Learn more about them (and exciting career opportunities) at: jane-st.co/computerphile
This video was filmed and edited by Sean Riley.
Computerphile is a sister project to Brady Haran's Numberphile. More at bradyharanblog.com










