How Google deals with AGIs potential threats to humanity @BuzzRobot
How Google deals with AGIs potential threats to humanity  @BuzzRobot
Uploaded May 2025 | Updated September 2026, 3 weeks ago
How does Google DeepMind treat AGI safety and the potential threats it poses to humanity as a whole?

In this video, @BuzzRobot talks to Rohin Shah, who leads the AGI Safety & Alignment team at Google DeepMind. Rohin shares the insights from a very extensive paper on AI safety measures, called "An Approach to Technical AGI Safety and Security". We talk about what gets identified as areas of severe risks, mitigation approaches, safety training and capability suppression, oversight, scaling AI safety measures, and more.

Read the full paper "An Approach to Technical AGI Safety and Security" here: arxiv.org/abs/2504.01849
Taking a responsible path to AGI blog post: https://deepmind.google/discover/blog/taking-a-responsible-path-to-agi/

If you'd like to catch these talks live and participate in Q&As with the speakers, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ

Timestamps:
0:00 Introduction
0:52 Context and Scope: Severe harm to humanity
3:10 Areas of severe risks
07:04 Technical approaches to misuse
10:34 Technical approaches to misalignment
15:11 Alignment approaches: Enablers
16:37 Safer design patterns

Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ

#ai #artificialintelligence #artificialgeneralintelligence #agi #machinelearning #deeplearning #neuralnetworks #llms #largelanguagemodels #deepmind #aisafety #googleai #technology #science #aitalk
How Google deals with AGIs potential threats to humanity4 skills to teach your kids in the #artificialgeneralintelligence worldAI agents trying to prevent their own shutdownAI that Helps Humans Find Common Ground in Controversial Issues. #largelanguagemodels  #airesearchCan you spot AI-generated content?In the race to create #artificialgeneralintelligence (AGI) safety took a back seat #machinelearningMost realistic path for AGI takeoverBlackmailing AI agents: are some models better than others?Soon there will be tons of developers writing shitty code. Heres why #artificialintelligence #aiAI builds its own tools to play Pokemon #aiDoes AI deserve compassion? AI welfare talkTraining Generative AI Models with Pixel Masking: Cropping vs Dropping #generativeai #pixelart #ai
BuzzRobot |

How Google deals with AGI's potential threats to humanity

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER