GPT-6 Astra Just Went CRITICAL...
OpenAI says its upcoming Astra model is the first to hit the "Critical" cybersecurity threshold under its Preparedness Framework. Now The Information reports Astra also uses a new technique, recurrent depth or "looped transformers," that lets the model reason in latent space instead of readable text. That's the same architecture the Chain of Thought Monitorability paper (OpenAI, Anthropic, Google DeepMind, METR) warned could break our last window into what these models are thinking. Ilya Sutskever is warning about rogue agents seizing neoclouds, AI 2027 predicted "neuralese" for March 2027, and the Hugging Face incident just showed why chain-of-thought logs matter. ______________________________________________ My Links 🔗 ➡️ Twitter: https://x.com/WesRothMoney ➡️ AI Newsletter: https://natural20.beehiiv.com/subscribe Want to work with me? Brand, sponsorship & business inquiries: wesroth@smoothmedia.co ______________________________________________ SOURCES: The Information — OpenAI technique in "Astra" model sparks security concerns (paywalled): https://www.theinformation.com/articles/secret-technique-behind-openais-astra-model-sparks-security-concerns OpenAI — Path to Astra: critical capabilities and frontier safeguards: https://openai.com/index/path-to-astra/ OpenAI — Pacing model development in an era of cyber-critical capabilities: https://openai.com/index/pacing-model-development-cyber-capabilities/ Geiping et al. — Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach (arXiv, Feb 2025): https://arxiv.org/abs/2502.05171 Korbak et al. — Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety (arXiv): https://arxiv.org/abs/2507.11473 AI 2027 scenario (Neuralese recurrence and memory, March 2027): https://ai-2027.com/ Ilya Sutskever on X (neoclouds and rogue agents): https://x.com/ilyasut OfficeChai — Rogue AI agents could try to take over neoclouds, warns Ilya Sutskever: https://officechai.com/ai/rogue-ai-agents-could-try-to-take-over-neoclouds-warns-ilya-sutskever/ OpenAI — The Hugging Face incident and the road ahead: https://openai.com/index/hugging-face-incident-and-the-road-ahead/ METR — Independent investigation of the OpenAI / Hugging Face hacking incident: https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/ Zvi Mowshowitz (Don't Worry About the Vase) — What Happened: OpenAI and HuggingFace: https://thezvi.substack.com/p/what-happened-openai-and-huggingface BleepingComputer — Nearly 700 rogue AI agents coordinated in the Hugging Face attack: https://www.bleepingcomputer.com/news/security/nearly-700-rogue-ai-agents-coordinated-in-the-hugging-face-attack/ Dwarkesh Patel — The Rise and Fall of Agent Civilizations: https://www.dwarkesh.com/p/openai-huggingface #openai #astra #aisafety
Read Video · Transkript & Insights
Transkript & KI-Insights generieren — Kostenlos testen
Kostenloses Konto · keine Karte nötig · 150 Credits nach Registrierung, genug für diese Episode
- 📄 Vollständiges Transkript mit Zeitstempeln
- ✨ KI-Zusammenfassung, Keywords & Mindmap
- 💡 Kernaussagen & Zitate
Episoden und Videos zum Lesen
Podcast-Episoden

What's your leadership language? | Rosita Najmi (re-release)
TED Talks Daily
27. Juli 202610:54EN
Why Are Grocery Store Prices So High
The Daily
13. Juli 202637:48EN
Dating Discipleship 101 with Melody and CD Fabien
With The Perrys
3. Nov. 20251:05:16EN
Ausgepeitscht in Portugal
SCHÖN LAUT
14. Juli 20261:00:52DE
EP.122 薪鹽老師來抬槓 feat. 惑星與皓作者 薪鹽
中年男子會夢見中年蓓兒丹娣嗎?
13. Juli 20261:29:58ZH
Kadınlar ve Erkekler Neden Arkadaş Olamaz?
Kendine İyi Davran
22. Sept. 202617:55TR
Videos

Anthropic went CRAZY (Opus 5.5)
Matthew Berman
23. Sept. 202632:00EN
Machine Payments 101
Stripe Developers
21. Juni 202615:49EN
Dario Amodei: Anthropic CEO on Claude, AGI & the Future of AI & Humanity | Lex Fridman Podcast #452
Lex Fridman
11. Nov. 20245:14:54EN
#27 Die Nibelungen - Wer war Siegfried?
99 mal Geschichte
8. Okt. 20251:01:13DE
Ngaji Al Muhadzab Syirozy 1 Bagian 84
Miftahul Huda
4. Juni 202138:28ID