Readpodcast AI

Inference for Async Agents in Production | Doubleword

AI Council
22/6/202624:07226 lượt xemXem trên YouTube

[2026 - Day 1 - AGENT INFRASTRUCTURE] As models improve, we are starting to build long-running, asynchronous agents such as deep research agents and browser agents that can execute multi-step workflows autonomously. These systems unlock new use cases, but they use orders of magnitude more tokens and compute, creating scaling bottlenecks. This talk discusses practical strategies builders can use to maximize async agent performance while keeping inference costs under control. Topics covered include context engineering, compaction, cache maintenance, model routing, and batch inference. This talk is aimed at use case developers, with secondary relevance to platform engineers. SPEAKER: Meryem Arik - Co-founder & CEO, Doubleword 👉 Sign up for our "No BS" Newsletter to get the latest technical data & AI content: https://aicouncil.com/newsletter ABOUT AI COUNCIL: AI Council brings together the brightest minds in data to share industry knowledge, technical architectures and best practices in building cutting edge data & AI systems and tools. FIND US: Website: https://aicouncil.com/ LinkedIn: https://www.linkedin.com/company/aicouncilconf/ X: https://x.com/aicouncilconf

Read Video · Bản Ghi & Phân Tích

Tạo bản ghi & AI insight — Dùng thử miễn phí

Tài khoản miễn phí · không cần thẻ · 150 tín dụng khi đăng ký, đủ để mở khóa tập này

  • 📄 Bản ghi đầy đủ có dấu thời gian
  • ✨ Tóm tắt AI, từ khóa & bản đồ tư duy
  • 💡 Điểm chính và trích dẫn nổi bật

Tập podcast và video sẵn sàng để đọc