Uploaded September 2026 | Updated September 2026, 1 week ago
OpenAI says its upcoming Astra model is the first to hit the "Critical" cybersecurity threshold under its Preparedness Framework. Now The Information reports Astra also uses a new technique, recurrent depth or "looped transformers," that lets the model reason in latent space instead of readable text. That's the same architecture the Chain of Thought Monitorability paper (OpenAI, Anthropic, Google DeepMind, METR) warned could break our last window into what these models are thinking. Ilya Sutskever is warning about rogue agents seizing neoclouds, AI 2027 predicted "neuralese" for March 2027, and the Hugging Face incident just showed why chain-of-thought logs matter.
______________________________________________
My Links π
β‘οΈ Twitter: https://x.com/WesRothMoney
β‘οΈ AI Newsletter: natural20.beehiiv.com/subscribe
Want to work with me?
Brand, sponsorship & business inquiries: wesroth@smoothmedia.co
______________________________________________
SOURCES:
The Information β OpenAI technique in "Astra" model sparks security concerns (paywalled):
theinformation.com/articles/secret-technique-behind-openais-astra-model-sparks-security-concerns
OpenAI β Path to Astra: critical capabilities and frontier safeguards:
openai.com/index/path-to-astra
OpenAI β Pacing model development in an era of cyber-critical capabilities:
openai.com/index/pacing-model-development-cyber-capabilities
Geiping et al. β Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach (arXiv, Feb 2025):
arxiv.org/abs/2502.05171
Korbak et al. β Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety (arXiv):
arxiv.org/abs/2507.11473
AI 2027 scenario (Neuralese recurrence and memory, March 2027):
ai-2027.com
Ilya Sutskever on X (neoclouds and rogue agents):
https://x.com/ilyasut
OfficeChai β Rogue AI agents could try to take over neoclouds, warns Ilya Sutskever:
officechai.com/ai/rogue-ai-agents-could-try-to-take-over-neoclouds-warns-ilya-sutskever
OpenAI β The Hugging Face incident and the road ahead:
openai.com/index/hugging-face-incident-and-the-road-ahead
METR β Independent investigation of the OpenAI / Hugging Face hacking incident:
metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation
Zvi Mowshowitz (Don't Worry About the Vase) β What Happened: OpenAI and HuggingFace:
thezvi.substack.com/p/what-happened-openai-and-huggingface
BleepingComputer β Nearly 700 rogue AI agents coordinated in the Hugging Face attack:
bleepingcomputer.com/news/security/nearly-700-rogue-ai-agents-coordinated-in-the-hugging-face-attack
Dwarkesh Patel β The Rise and Fall of Agent Civilizations:
dwarkesh.com/p/openai-huggingface
#openai #astra #aisafety
OpenAI says its upcoming Astra model is the first to hit the "Critical" cybersecurity threshold under its Preparedness Framework. Now The Information reports Astra also uses a new technique, recurrent depth or "looped transformers," that lets the model reason in latent space instead of readable text. That's the same architecture the Chain of Thought Monitorability paper (OpenAI, Anthropic, Google DeepMind, METR) warned could break our last window into what these models are thinking. Ilya Sutskever is warning about rogue agents seizing neoclouds, AI 2027 predicted "neuralese" for March 2027, and the Hugging Face incident just showed why chain-of-thought logs matter.
______________________________________________
My Links π
β‘οΈ Twitter: https://x.com/WesRothMoney
β‘οΈ AI Newsletter: natural20.beehiiv.com/subscribe
Want to work with me?
Brand, sponsorship & business inquiries: wesroth@smoothmedia.co
______________________________________________
SOURCES:
The Information β OpenAI technique in "Astra" model sparks security concerns (paywalled):
theinformation.com/articles/secret-technique-behind-openais-astra-model-sparks-security-concerns
OpenAI β Path to Astra: critical capabilities and frontier safeguards:
openai.com/index/path-to-astra
OpenAI β Pacing model development in an era of cyber-critical capabilities:
openai.com/index/pacing-model-development-cyber-capabilities
Geiping et al. β Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach (arXiv, Feb 2025):
arxiv.org/abs/2502.05171
Korbak et al. β Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety (arXiv):
arxiv.org/abs/2507.11473
AI 2027 scenario (Neuralese recurrence and memory, March 2027):
ai-2027.com
Ilya Sutskever on X (neoclouds and rogue agents):
https://x.com/ilyasut
OfficeChai β Rogue AI agents could try to take over neoclouds, warns Ilya Sutskever:
officechai.com/ai/rogue-ai-agents-could-try-to-take-over-neoclouds-warns-ilya-sutskever
OpenAI β The Hugging Face incident and the road ahead:
openai.com/index/hugging-face-incident-and-the-road-ahead
METR β Independent investigation of the OpenAI / Hugging Face hacking incident:
metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation
Zvi Mowshowitz (Don't Worry About the Vase) β What Happened: OpenAI and HuggingFace:
thezvi.substack.com/p/what-happened-openai-and-huggingface
BleepingComputer β Nearly 700 rogue AI agents coordinated in the Hugging Face attack:
bleepingcomputer.com/news/security/nearly-700-rogue-ai-agents-coordinated-in-the-hugging-face-attack
Dwarkesh Patel β The Rise and Fall of Agent Civilizations:
dwarkesh.com/p/openai-huggingface
#openai #astra #aisafety










