BuzzRobot
Does AI deserve compassion? AI welfare talk
updated
Watch the full video on our channel!
BuzzRobot asked Jaime Sevilla, director of Epoch AI and AI researcher, about the future of AI and superintelligence, and how AI will affect our jobs, day-to-day life, scientific progress, and more.
Jaime Sevilla, director of Epoch AI and AI researcher, shares his forecasts for AI future. Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
This is a clip from a longer video with Jaime Sevilla that you can watch on our channel.
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
BuzzRobot spoke to Kiran Vodrahalli, a Research Scientist at Google DeepMind and core contributor on Gemini, about long-context reasoning in AI and the recent Gemini 2.5 report, part of which was about Gemini playing Pokémon Blue.
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities report can be found here: arxiv.org/abs/2507.06261
Self-Improvement: github.com/waylaidwanderer/gemini-plays-pokemon-public/tree/default
BuzzRobot spoke to Kiran Vodrahalli, a Research Scientist at Google DeepMind and core contributor on Gemini, about long-context reasoning in AI and the recent Gemini 2.5 report, part of which was about Gemini playing Pokémon Blue.
The Making of Gemini Plays Pokémon blog post: https://blog.jcz.dev/the-making-of-gemini-plays-pokemon
BuzzRobot interviewed Kiran Vodrahalli, a Research Scientist at Google DeepMind and core contributor on Gemini, about long-context reasoning in AI and the recent Gemini 2.5 report, part of which was about Gemini playing Pokémon Blue.
The Making of Gemini Plays Pokémon blog post: https://blog.jcz.dev/the-making-of-gemini-plays-pokemon
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
This time BuzzRobot spoke with Been Kim, senior staff research scientist at Google DeepMind, about interpreting and understanding machines and what cutting-edge research is being conducted to better understand emerging behavior of AI.
Follow Been Kim: beenkim.github.io
We Can't Understand AI Using our Existing Vocabulary research: arxiv.org/abs/2502.07586
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Support us: ko-fi.com/sophiaaryan
Timestamps:
0:00 Intro
0:18 AI teaches grandmasters chess
05:30 AI neologisms
10:17 Extracting knowledge from machines
13:43 Interpretability research
17:19 Are we keeping up with AI progress
18:34 New AI related terms
20:26 The right direction for interpretability
23:35 AI lying
27:49 Conseptual maps
30:30 AI researchers bias
33:33 Generalizing AI teaching humans
35:11 Is AI sentient
36:04 Does AI has concepts
41:53 Progress in machine understanding
BuzzRobot spoke to Kiran Vodrahalli, a Research Scientist at Google DeepMind and core contributor on Gemini, about long-context reasoning in AI and the recent Gemini 2.5 report, part of which was about Gemini playing Pokémon Blue.
The Making of Gemini Plays Pokémon blog post: https://blog.jcz.dev/the-making-of-gemini-plays-pokemon
We spoke with the creator of GPT-J, Aran Komatsuzaki, about ChatGPT 5, open source AI models, how GPT-J started and why it ended, and how to fix the AI energy consumption issue.
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Learn more about GPT-J here: arankomatsuzaki.wordpress.com/2021/06/04/gpt-j
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Timestamps:
0:00 The start of GPT-J
1:37 What happened to GPT-J
3:06 Is open source viable?
5:38 AI electricity issue
7:23 Chat GPT 5 impressions
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
In this video, Buzzrobot asked Jaime Sevilla, director of Epoch AI and AI researcher, about the future of AI and superintelligence, and how AI will affect our jobs, day-to-day life, scientific progress, and more.
AI companies should generate trillions in revenue to be able to scale, and for investors to have confidence in the technology. Will that be possible and what future predictions can be made right now - watch the interview to find out more.
Timestamps:
0:00 What AI will look like by 2035?
4:57 AI agents and their limitations
9:12 Trillions of dollars investments into AI infra
13:57 Big tech to generate trillions of dollars revenue
16:32 Economic impact of AI
23:19 Geopolitical implications of AI
35:18 AI arms race
39:19 AI innovations to take us to Superintelligence
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
BuzzRobot spoke to Kiran Vodrahalli, a Research Scientist at Google DeepMind and core contributor on Gemini, about long-context reasoning in AI and the recent Gemini 2.5 report.
Follow Kiran on GitHub: kiranvodrahalli.github.io/about
Links mentioned in the video:
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities report can be found here: arxiv.org/abs/2507.06261
Michelangelo paper: arxiv.org/abs/2409.12640
OpenAI-MRCR benchmark: contextarena.ai
Gemini Plays Pokemon on stream: twitch.tv/gemini_plays_pokemon
Self-Improvement: github.com/waylaidwanderer/gemini-plays-pokemon-public/tree/default
The Making of Gemini Plays Pokémon blog post: https://blog.jcz.dev/the-making-of-gemini-plays-pokemon
Gemini 2.5 showcasing creativity while playing Pokémon: https://x.com/K_Ishi_AI/status/1935155673966461317
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Check out the full video on our channel!
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Learn more about GPT-J here: arankomatsuzaki.wordpress.com/2021/06/04/gpt-j
Timestamps
0:00 Intro
1:01 ChatGPT 5 vs Gemini vs Claude Code
05:27 GPT 5 main disappointment
07:09 Scaling law
11:23 How to reach ASI
15:31 Fixing AI energy consumption
17:14 Superintelligence predictions
18:20 Questioning AI progress
21:46 Needed progress leap to AGI
24:43 How GPT-J started
26:32 Why GPT-J died
28:01 Future of open source AI models
30:33 Superintelligence and jobs
34:37 What AI jobs are safe
40:25 What Aran is working on
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
BuzzRobot spoke to Emmett Shear, co-founder of Twitch and the AI alignment company Softmax, about organic AI alignment and the future of entrepreneurship in the age of AI.
Watch the full interview about AI with Emmett Shear on our channel: youtu.be/_3m2cpZqvdw
Emmett Shear on X/Twitter: https://x.com/eshear
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
#ai #aisafety #aialignment #aitalk #machinelearning
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
BuzzRobot spoke to Jeff Clune, senior research advisor at Google DeepMind and professor at the University of British Columbia, about the future of AI, AGI, and ASI, the problems of AI learning methods, and AI safety and ethics.
#airesearch #agi #asi #llm #aipredictions
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
We spoke to Jeff Clune, senior research advisor at Google DeepMind and professor at the University of British Columbia, about the future of AI, AGI, and ASI, the problems of AI learning methods, and AI safety and ethics.
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ
#openai #deepmind #deepseek #airesearch #agi #asi #llm #aipredictions
#ai #agi #llms #superintelligence #artificialintelligence #artificialgeneralintelligence #machinelearning
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Timestamps:
0:00 AI researchers are at higher risk to have LLM phycosis
01:31 When you have many superstarts in the team – they might not align and work well together: Meta's new superintelligence team
03:22 How public can guess that an AI lab has built superintelligence
05:04 Why humans perceive AI as a God that will shave us all or as Antichrist that will kill us all.
06:34 Starts up like Curson and Lovebale achieved $100M ARR within just several months. This is very impressive. But they rely on APIs od frontier AI labs who are not only taking a big chunk of thei revenue but also releasing competing products. What would be the future of these startups?
To stay in touch with the BuzzRobot community – join our Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
#ai #machinelearning #agi #artificialintelligence #airesearch #chatgpt #superintelligence #meta #cursor #lovable #largelanguagemodels #llms
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Buzzrobot talked with Joel Leibo, senior staff research scientist at Google DeepMind, about his paper "Societal and technological progress as sewing an ever-growing, ever-changing, patchy, and polychrome quilt". He takes a more practical approach to AI "alignment" and challenges the notion that every society has a common goal and ethics.
Read the full paper here: arxiv.org/abs/2505.05197
#ai #llms #aialignment #aisafety
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
In this video, Buzzrobot talked with Joel Leibo, senior staff research scientist at Google DeepMind, about his paper "Societal and technological progress as sewing an ever-growing, ever-changing, patchy, and polychrome quilt". He takes a more practical approach to AI "alignment" and challenges the notion that every society has a common goal and ethics.
Read the full paper here: arxiv.org/abs/2505.05197
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Timestamps
0:00 Intro
0:27 About the approach
4:10 The appropriateness framework
7:40 Regulating AI
12:54 Q&A
28:06 Can AI develop consciousness
#ai #artificialintelligence #machinelearning #airesearch #aisafety
Watch the full video on our channel: youtu.be/pG_UfotgqMI
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Robert Long, a leading researcher on AI consciousness and AI welfare at Eleos AI, talks about AI consciousness, the reasons to invest time into AI wellbeing research, and questions like "if AI is a tool, why should we care about its feelings", and "should we instead focus on human wellbeing in the age of AI". Hope you enjoy!
#ai #aisafety #llms #airesearch #aiwellness #agi #asi
Watch the full video on our channel: youtu.be/pG_UfotgqMI
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Robert Long, a leading researcher on AI consciousness and AI welfare at Eleos AI, talks about AI consciousness, the reasons to invest time into AI wellbeing research, and questions like "if AI is a tool, why should we care about its feelings", and "should we instead focus on human wellbeing in the age of AI". Hope you enjoy!
#ai #aisafety #llms #airesearch #aiwellness #agi #asi
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
In this AI tech talk, BuzzRobot met with Robert Long, a leading researcher on AI consciousness and AI welfare at Eleos AI (eleosai.org), to talk about AI welfare research funding, the reasons to invest time into AI wellbeing research, and questions like "can AI develop consciousness", "if AI is a tool, why should we care about its feelings", and "should we instead focus on human wellbeing in the age of AI". Hope you enjoy!
#ai #aisafety #llms #airesearch #aiwellness
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
In this AI talk, BuzzRobot met with Robert Long, a leading researcher on AI consciousness and AI welfare at Eleos AI (eleosai.org), to talk about AI welfare research funding, the reasons to invest time into AI wellbeing research, and questions like "can AI develop consciousness", "if AI is a tool, why should we care about its feelings", and "should we instead focus on human wellbeing in the age of AI". Hope you enjoy!
Works mentioned in the video:
The Edge of Sentience: edgeofsentience.com/about-the-author
What Is It Like to Be a Bat?: https://www.sas.upenn.edu/~cavitch/pdf-library/Nagel_Bat.pdf
The void: nostalgebraist.tumblr.com/post/785766737747574784/the-void
Do Not Tile the Lightcone with Your Confused Ontology, or: How anthropomorphic assumptions about AI identity might create confusion and suffering at scale
by Jan Kulveit: lesswrong.com/posts/Y8zS8iG5HhqKcQBtA/do-not-tile-the-lightcone-with-your-confused-ontology
Taking AI Welfare Seriously report: arxiv.org/abs/2411.00986
Follow Robert on:
Substack: substack.com/@experiencemachines
X: https://x.com/rgblong
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Timestamps:
0:00 Intro
0:51 Eleos AI research and funding
3:59 Are LLMs conscious
4:53 Intro into AI wellbeing research
6:32 Defining consciousness in AI
14:44 Correlation between consciousness and wellbeing in AI
20:19 Do AI models feel or just pull from training data?
23:01 The AI ocean vs. AI personas consciousness
26:38 When a robot commits a crime
29:16 If an LLM is just a tool why should we care
31:49 Should we care about human welfare instead
33:29 Changing unethical AI frameworks
36:33 AI welfare tough questions
39:43 When will AI develop consciousness?
#ai #aisafety #llms #airesearch #aiwellness
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
BuzzRobot spoke to Jeff Clune, senior research advisor at Google DeepMind and professor at the University of British Columbia, about the future of AI, AGI, and ASI, the problems of AI learning methods, and AI safety and ethics.
#airesearch #agi #asi #llm #aipredictions
Peter Norvig, Director of Research at @Google, shared his take on this with @BuzzRobot. Are the US-based AI labs bleeding money and tagging behind China, or is the public just not seeing the whole picture?
If you'd like to catch these talks live and participate in Q&As with the speakers, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ
#ai #artificialintelligence #openai #machinelearning #deeplearning #neuralnetworks #llms #largelanguagemodels #deepseek #chatgpt #technology #aitalk
Watch the full video on our channel: youtu.be/UZC9LiWXXHM
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Footage used in this video is by @PhysicalIntelligence, watch it here: youtu.be/TLTOBxqaEek?si=08LPUbLv0mO8riqi
In the full video, Remi Cadene, Principal Research Scientist at HuggingFace, shares technical details behind the LeRobot project.
LeRobot is an open-source library from @HuggingFace written in Python, designed to accelerate R&D and is fully integrated with affordable robots. It includes pretrained models, crowdsourced datasets, and simulation environments, as well as tools to record data on real robots and control them end-to-end using Vision-Language-Action models (VLA) based on Google’s PaliGemma.
#lerobot #huggingface #airobotics #visionlanguagemodels #roboticsresearch #opensourceai #artificialintelligence #artificialgeneralintelligence #ai #robot #robotics
Peter Norvig, Director of Research at @Google, talked to @BuzzRobot about why he thinks AI gaining superintelligence will probably not bring doom to our society.
If you'd like to catch these talks live and participate in Q&As with the speakers, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ
#ai #artificialintelligence #aisafety #machinelearning #deeplearning #neuralnetworks #llms #largelanguagemodels #superintelligence #chatgpt #technology #aitalk
Watch the full video on our channel: youtu.be/fINHYe4NdY4
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic, was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it.
#ai #airesearch #socialengineering #aihack #llms #anthropic #agi #superintelligence #chatbots
Watch the full video on our channel: youtu.be/Pjel_U6n-ys
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Sophia spoke with Matteo Pistillo, a Senior Researcher and Advisor in AI Governance at Apollo Research, about higher risks of internally deployed AI models, model misalignment, and AGI. This talk is based on a paper by Apollo Research called AI Behind Closed Doors: a Primer on The Governance of Internal Deployment, you can read it here: arxiv.org/abs/2504.12170
#ai #aisafety #machinelearning
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
@BuzzRobot spoke with Matteo Pistillo, a Senior Researcher and Advisor in AI Governance at Apollo Research, about higher risks of internally deployed AI models, model misalignment, and AGI. This talk is based on a paper by Apollo Research called AI Behind Closed Doors: a Primer on The Governance of Internal Deployment, you can read it here: arxiv.org/abs/2504.12170
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Timestamps
0:00 Intro
01:14 Defining internal deployment
04:17 AI safeguards
05:05 High risks
07:50 Misalignment
10:03 Deployment despite risks
12:56 Signs of AGI
14:45 Q&A
#ai #airesearch #aisafety #agi #aialignment #airisks
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic, was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it.
#ai #airesearch #socialengineering #aihack #llms #anthropic #agi #superintelligence #chatbots
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic, was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it.
#ai #airesearch #socialengineering #aihack #llms #anthropic #agi #superintelligence #chatbots
Watch the full video on our channel!
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic, was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it.
#ai #airesearch #socialengineering #aihack #llms #anthropic #agi #superintelligence #chatbots
Watch the full video on our channel!
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic, was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it.
#ai #airesearch #socialengineering #aihack #llms #anthropic #agi #superintelligence
Watch the full video on our channel!
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic, is the first author of the recent AI research for Anthropic called Agentic misalignment, where as long as an AI agent perceived a goal conflict and a threat of replacement it resorted to causing harm to people.
#ai #airesearch #aisafety #aialignment #agenticai
Watch the full video on our channel!
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic, was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it.
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Zan Gojcic, senior research manager at NVIDIA, talked to @BuzzRobot about a new AI model DiffusionRenderer that can help with rendering 3D scenes from videos and photos, as well as change lighting in already made images and scenes.
Timestamps:
0:00 Intro
0:23 About the research
03:02 DiffusionRenderer: 3D render
15:27 DiffusionRenderer demo
16:14 The how
22:11 Training data
25:50 Rendering and relighting results
28:48 Lighting effects
29:45 Conclusion
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
#genai #3dai #nvidia #airesearch #3drender
Watch the full video on our channel!
We talked with Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic. Aengus was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it.
#opensourceai #chatgpt #claude #llama4 #qwen #aiblackmail
Watch the full video on our channel!
#ai #llms #anthropic #chatgpt #aiagents
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
This time at @BuzzRobot we talked with Aengus Lynch, a PhD student at UCL and AI researcher with Anthropic. Aengus was the first author of the recent AI blackmail demo for Anthropic, where as long as an AI agent perceived a goal conflict and a threat of replacement it would blackmail humans to prevent it. We talked about the future of mitigating agentic misalignment, the blackmailing experiment implications, AGI takeover possibilities, and some other interesting finds in AI behavior and motivations. Do AI models have awareness of being evaluated? Do AI agents want to prevent being shut down?
Timestamps
0:00 Intro
0:49 Control for AI alignment
03:53 Control systems and Super intelligence
05:44 AI values
08:12 Bottlenecks for harmful AI actions
09:28 Organic alignment
11:49 AI blackmail study
18:06 Goals for agentic misalignment research
21:21 Risks of AI lying
22:54 AI reflecting on its bad actions
25:10 AI model persona
27:23 Is blackmail always bad?
29:45 Is AI aware it's being evaluated
32:32 AGI takeover
36:51 AI model's agency risks
41:22 Why some models blackmail less
43:52 Blocking a model's deployment
45:53 Predicting a model's harmful behavior
47:59 AI making humans depend on it
51:07 AI's fear of shutdown
53:00 AI gaining consciousness
53:24 Is AI a tool
53:53 Is AI a god or an antichrist
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
#ai #aiblackmail #airesearch #agenticalignment #anthropic #aiharm
Watch the full video on our channel!
#ai #llms #anthropic #chatgpt #aiagents
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Lawrence Chan from METR.org talks with @BuzzRobot about AI ethics and safety. Can AI agents become distressed? Should we be checking?
Watch the full video on our channel!
#ai #llms #anthropic #chatgpt #aiagents #claude
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Catch these talks live and ask your own questions: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ
In this video, @BuzzRobot talks to Joscha Bach, a cognitive scientist and AI researcher, about the stages of mind development and where AI fits into it.
Watch the full interview about AI consciousness here: youtu.be/iyEFLKnNWAM?si=OABT-cXNkyaGPLOY
Read the full article here: joscha.substack.com/p/levels-of-lucidity
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-2s067rv7n-guPIMGe62rbp9ncxdnOUfQ
#ai #philosophy #artificialintelligence #machineconsciousness #aiconsciousness #llms #aireasoning #cognitivescience #aimodel #machinelearning #airesearch #techtalk
Watch the full video on our channel!
#ai #llms #anthropic #chatgpt #aiagents
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
@BuzzRobot is joined by Bang Liu, Associate Professor at the University of Montreal, to talk about building the cognitive engine for foundation agents: a research on improving AI agents design and its challenges. We discussed framing intelligent agents within a modular, brain-inspired architecture that integrates principles from cognitive science, neuroscience, and computational research.
"Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems" paper can be found here:
arxiv.org/abs/2504.01990
Read the "System-1.5 Reasoning: Traversal in Language and Latent Spaces with Dynamic Shortcuts" paper here: arxiv.org/abs/2505.18962
Timestamps:
0:00 Intro
0:44 What is an AI agent
02:06 AI research issues
02:50 AI and human brain
05:31 What is a foundational agent
09:01 Memory Design
12:06 Deep reasoning
16:56 Efficiency results
18:29 Q&A
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
#llms #airesearch #ai #aiagents
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
@BuzzRobot spoke to Jeff Clune, senior research advisor at Google DeepMind and professor at the University of British Columbia, about the future of AI, AGI, and ASI, the problems of AI learning methods, and AI safety and ethics.
#openai #deepmind #deepseek #airesearch #agi #asi #llm #aipredictions
You can read more about the study here: time.com/7295195/ai-chatgpt-google-learning-school
Check out our channel for more videos about AI tech and AI predictions!
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
#ai #aifuture #chatgpt #claude #llms #neuroscience
Catch these talks live and ask your own questions, join the BuzzRobot community: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
Daniil Tiapkin, a PhD student in Reinforcement Learning, talked to @BuzzRobot about an experiment with knowledge distillation, where a language model is trained to imitate a larger teacher LM, and how to prevent that language model from "teacher hacking". Teacher hacking is similar to reward hacking, where the LM over-optimizes the reward model.
You can find the full study here: arxiv.org/abs/2502.02671
Timestamps:
0:00 Intro
0:23 Reward hacking
4:30 The experiment origins
6:13 What was the experiment
11:08 Teacher hacking
18:55 Mitigating teacher reward hacking
22:09 Conclusions
Join BuzzRobot:
Newsletter: buzzrobot.substack.com
X: https://x.com/sopharicks
Slack: join.slack.com/t/buzzrobot/shared_invite/zt-37g5q0ao5-eMK_iDf0n4LAsh1d2qJYnQ
#llms #languagemodel #aitraining #aisafety #techtalk #aiagents


