GPT-6 Astra Just Went CRITICAL...
OpenAI says its upcoming Astra model is the first to hit the "Critical" cybersecurity threshold under its Preparedness Framework. Now The Information reports Astra also uses a new technique, recurrent depth or "looped transformers," that lets the model reason in latent space instead of readable text. That's the same architecture the Chain of Thought Monitorability paper (OpenAI, Anthropic, Google DeepMind, METR) warned could break our last window into what these models are thinking. Ilya Sutskever is warning about rogue agents seizing neoclouds, AI 2027 predicted "neuralese" for March 2027, and the Hugging Face incident just showed why chain-of-thought logs matter. ______________________________________________ My Links 🔗 ➡️ Twitter: https://x.com/WesRothMoney ➡️ AI Newsletter: https://natural20.beehiiv.com/subscribe Want to work with me? Brand, sponsorship & business inquiries: wesroth@smoothmedia.co ______________________________________________ SOURCES: The Information — OpenAI technique in "Astra" model sparks security concerns (paywalled): https://www.theinformation.com/articles/secret-technique-behind-openais-astra-model-sparks-security-concerns OpenAI — Path to Astra: critical capabilities and frontier safeguards: https://openai.com/index/path-to-astra/ OpenAI — Pacing model development in an era of cyber-critical capabilities: https://openai.com/index/pacing-model-development-cyber-capabilities/ Geiping et al. — Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach (arXiv, Feb 2025): https://arxiv.org/abs/2502.05171 Korbak et al. — Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety (arXiv): https://arxiv.org/abs/2507.11473 AI 2027 scenario (Neuralese recurrence and memory, March 2027): https://ai-2027.com/ Ilya Sutskever on X (neoclouds and rogue agents): https://x.com/ilyasut OfficeChai — Rogue AI agents could try to take over neoclouds, warns Ilya Sutskever: https://officechai.com/ai/rogue-ai-agents-could-try-to-take-over-neoclouds-warns-ilya-sutskever/ OpenAI — The Hugging Face incident and the road ahead: https://openai.com/index/hugging-face-incident-and-the-road-ahead/ METR — Independent investigation of the OpenAI / Hugging Face hacking incident: https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/ Zvi Mowshowitz (Don't Worry About the Vase) — What Happened: OpenAI and HuggingFace: https://thezvi.substack.com/p/what-happened-openai-and-huggingface BleepingComputer — Nearly 700 rogue AI agents coordinated in the Hugging Face attack: https://www.bleepingcomputer.com/news/security/nearly-700-rogue-ai-agents-coordinated-in-the-hugging-face-attack/ Dwarkesh Patel — The Rise and Fall of Agent Civilizations: https://www.dwarkesh.com/p/openai-huggingface #openai #astra #aisafety
Read Video · 文字稿与深度分析
一键获取文字稿与 AI 深度分析 — 免费体验
免费注册 · 无需信用卡 · 注册即获 150 积分,足够解锁本集
- 📄 完整文字稿含时间戳
- ✨ AI 摘要、关键词与思维导图
- 💡 核心要点与精彩引言
播客与视频,已可阅读
音频播客

The Unprecedented Personal Profits of Trump’s Presidency
The Daily
2026年7月9日29:56EN
What a harness is and how to build one with Claude Agent SDK
How I AI
2026年7月8日24:35EN
Miele’s Andreas Wieser and Yulia Kalner at Kantar on reinventing German Brands Through Emotion
DMEXCO Podcast by Verena Gründel
2026年7月29日31:55EN
假期通知兼谈本台为什么要做视频播客
商业就是这样
2026年9月27日9:12ZH-Hans
#24 「やるべきは明石家さんまだった」
朝井リョウ・加藤千恵 信頼できない語り手
2026年7月17日52:14JA
#357 HATZE
Gemischtes Hack
2026年8月11日1:30:28DE
视频

Kate McAndrew: The $100M Investing Playbook l Trailblazers Podcast Episode 40
Trailblazers with Erica Wenger
2026年7月8日1:12:18EN
What is Loop Engineering?
KodeKloud
2026年7月14日6:47EN
Brad Gerstner: No AI Bubble, Semis Eat the Nasdaq & AI's Take Off Problem
All-In Podcast
2026年9月17日18:10EN
从「上瘾模型」到「专注力训练」,如何在被算法理解的世界里重新找回主动?| 英文访谈 S9E33
声动活泼
2025年10月16日50:48ZH-Hans
Firma bez šéfov: Funguje to? - Money Talk 113 s Ferom Baníkom
Milan Dubec
2026年8月4日54:10SK