Readpodcast AI

GPT-6 Astra Just Went CRITICAL...

Wes Roth
2026/9/218:58236,567 次觀看在 YouTube 觀看

OpenAI says its upcoming Astra model is the first to hit the "Critical" cybersecurity threshold under its Preparedness Framework. Now The Information reports Astra also uses a new technique, recurrent depth or "looped transformers," that lets the model reason in latent space instead of readable text. That's the same architecture the Chain of Thought Monitorability paper (OpenAI, Anthropic, Google DeepMind, METR) warned could break our last window into what these models are thinking. Ilya Sutskever is warning about rogue agents seizing neoclouds, AI 2027 predicted "neuralese" for March 2027, and the Hugging Face incident just showed why chain-of-thought logs matter. ______________________________________________ My Links 🔗 ➡️ Twitter: https://x.com/WesRothMoney ➡️ AI Newsletter: https://natural20.beehiiv.com/subscribe Want to work with me? Brand, sponsorship & business inquiries: wesroth@smoothmedia.co ______________________________________________ SOURCES: The Information — OpenAI technique in "Astra" model sparks security concerns (paywalled): https://www.theinformation.com/articles/secret-technique-behind-openais-astra-model-sparks-security-concerns OpenAI — Path to Astra: critical capabilities and frontier safeguards: https://openai.com/index/path-to-astra/ OpenAI — Pacing model development in an era of cyber-critical capabilities: https://openai.com/index/pacing-model-development-cyber-capabilities/ Geiping et al. — Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach (arXiv, Feb 2025): https://arxiv.org/abs/2502.05171 Korbak et al. — Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety (arXiv): https://arxiv.org/abs/2507.11473 AI 2027 scenario (Neuralese recurrence and memory, March 2027): https://ai-2027.com/ Ilya Sutskever on X (neoclouds and rogue agents): https://x.com/ilyasut OfficeChai — Rogue AI agents could try to take over neoclouds, warns Ilya Sutskever: https://officechai.com/ai/rogue-ai-agents-could-try-to-take-over-neoclouds-warns-ilya-sutskever/ OpenAI — The Hugging Face incident and the road ahead: https://openai.com/index/hugging-face-incident-and-the-road-ahead/ METR — Independent investigation of the OpenAI / Hugging Face hacking incident: https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/ Zvi Mowshowitz (Don't Worry About the Vase) — What Happened: OpenAI and HuggingFace: https://thezvi.substack.com/p/what-happened-openai-and-huggingface BleepingComputer — Nearly 700 rogue AI agents coordinated in the Hugging Face attack: https://www.bleepingcomputer.com/news/security/nearly-700-rogue-ai-agents-coordinated-in-the-hugging-face-attack/ Dwarkesh Patel — The Rise and Fall of Agent Civilizations: https://www.dwarkesh.com/p/openai-huggingface #openai #astra #aisafety

Read Video · 文字稿與深度分析

一鍵獲取文字稿與 AI 深度分析 — 免費體驗

免費註冊 · 無需信用卡 · 註冊即獲 150 積分,足夠解鎖本集

  • 📄 完整文字稿含時間戳
  • ✨ AI 摘要、關鍵詞與心智圖
  • 💡 核心要點與精彩引言

Podcast 與影片,已可閱讀