GPT-6 SOL AND LUNA ARE OUT!!!
Join My Newsletter for Regular AI Updates 👇🏼 https://forwardfuture.com My Links 🔗 👉🏻 X: https://x.com/matthewberman 👉🏻 Forward Future X: https://x.com/forwardfuture 👉🏻 Instagram: https://www.instagram.com/matthewberman_ai 👉🏻 Discord: https://discord.gg/u7wTTGWhuJ 👉🏻 Spotify: https://open.spotify.com/show/6dBxDwxtHl1hpqHhfoXmy8 Media/Sponsorship Inquiries ✅ https://bit.ly/44TC45V
동영상 읽기 · 대본 및 인사이트
이 에피소드에는 전체 대본과 AI 인사이트가 있습니다
무료 계정 · 카드 불필요 · 가입 시 크레딧 150개, 이 에피소드를 잠금 해제하기에 충분합니다
- 📄 타임스탬프가 포함된 전체 대본
- ✨ AI 요약, 키워드와 마인드맵
- 💡 주요 결론과 인용문
- 화자 1
에피소드 타임라인
Introduction and overview of Claude Opus 5.5 release
- Anthropic released Claude Opus 5.5, the first model since Dario's essay calling for pacing the frontier, and it performs at Fable 5.1 level for most tasks at 40% lower cost than Opus 5.
- Opus 5.5 is faster, cheaper, and on many benchmarks actually takes the frontier position from Fable 5.1, which the host calls crazy to see.
- The host notes this is the first major model release since Anthropic began pacing the frontier, and external evaluators Meter and Frontier Design tested it before release.
Benchmark deep dive: Terminal Bench, Frontier Code, Cursor Bench, GDPval
- Opus 5.5 dominates Terminal Bench 4.0 with 66.4% versus Astra's 57.9% and Fable 5.1's 55.8%, a 10-point jump over Fable.
- On GDPval 2.1, OpenAI's benchmark for real-world knowledge work, Opus 5.5 scores 1846 ELO versus Fable 5.1's 1735, a 300+ point jump over GPT6 Astra's 1542.
- On Frontier Code V1.1 Opus 5.5 hits 54.4% versus Fable's 50% and Astra's 53.3%, and on Cursor Bench it scores 57.8% versus Fable 5.1's 51%.
- Opus 5.5 did not take first place on Automation Bench (Astra 41.4% vs 40%) or Terminal Bench Science, where GPT6 Astra still leads.
Pricing, speed, and cost-per-task efficiency analysis
- Opus 5.5 costs $4 per million input tokens and $20 per million output tokens, a 20% price decrease from Opus 5, with cache reads at 20 cents versus 50 cents.
- The host argues cost per task completed matters more than cost per token, since a cheap model that uses 10x tokens is still expensive overall.
- On Automation Bench and Frontier Code, Opus 5.5 sits in the top-left quadrant of quality versus cost, with medium effort beating max effort on cost efficiency.
- Opus 5.5 generates output more than 30% faster than Opus 5 and requires less compute to serve.
아직 생성된 콘텐츠가 없습니다.
아직 생성된 콘텐츠가 없습니다.
아직 생성된 콘텐츠가 없습니다.
대본과 인사이트는 AI가 생성하므로 오류가 있을 수 있습니다. 정확도는 오디오 품질과 화자의 명료도에 따라 달라지며, 이상한 부분이 있다면 원본 오디오가 항상 기준입니다.
바로 읽을 수 있는 에피소드와 동영상
팟캐스트 에피소드
INTJ Personality Type Advice - 0088
Personality Hacker Podcast
2015년 10월 19일1:05:25EN
A journalist's trick for talking to people you can't stand | Joshua Johnson
TED Talks Daily
2026년 8월 3일9:31EN
Humanity's First Star Probe, Architect Labs Beats NVIDIA 3.4x, Musk Wants Satellites to Cool Earth | EP #285
Moonshots with Peter Diamandis
2026년 9월 2일1:57:22EN
Vol.257 为什么上证是“综指”、深证用“成指”?
商业就是这样
2026년 5월 20일29:12ZH
Sind Deutschland Kinder egal? Mit Caroline von St. Ange
Politik mit Anne Will
2026년 5월 22일2:01:19DE
SÉRIE: RELIGIÃO TÓXICA - A GRAÇA NÃO É O QUE VOCÊ PENSA| PR.YAN AUGUSTO
Minha Igreja Na Cidade
2026년 7월 27일55:27PT
동영상

Dario Amodei — “We are near the end of the exponential”
Dwarkesh Patel
2026년 2월 13일2:22:01EN
Anthropic's CEO: ‘We Don’t Know if the Models Are Conscious’ | Interesting Times with Ross Douthat
Interesting Times
2026년 2월 12일1:02:30EN
Staphylococcus aureus
Osmosis from Elsevier
2020년 10월 14일14:32EN
#32 Die Schlacht von Worringen - Der Freiheitskampf der Kölner
99 mal Geschichte
2025년 11월 12일53:29DE
10 habitudes qui m’ont VRAIMENT fait perdre du poids
leawellnesss
2026년 4월 15일14:34FR