From Spans to Trajectories: Observability for Long-Running Agents | HoneyHive
[2026 - DAY 2 - AI ENGINEERING] Agents have evolved. We've moved from orchestration frameworks — where agents operate through definite steps and turns you defined upfront — to harnesses, where the LLM uses skills and tools to chart its own trajectory. Modern agents run for hours or days, producing hundreds to thousands of steps in a single session. This calls for a fundamentally different methodology for monitoring and evaluating them in production. This talk shares what we've learned building observability infrastructure for agent harnesses at HoneyHive. We'll start with why the harness — not the model — has become the hardest engineering problem in production AI, and why traditional APM breaks down when traces are 10,000 spans deep and failures happen four tool calls deep. We'll walk through a live trajectory view to see what long-running agent traces actually look like at scale, and the specific challenges they create: context rot, semantic failure modes, and the needle-in-a-haystack problem of finding the moment that mattered. Then we'll dig into skills as the new unit of behavior and the dual role of clustering in agent development: unsupervised clustering for discovering emergent patterns and identifying where guardrails are needed, and supervised classifiers for production evaluation at scale. We'll close on what comes next — swarm observability for multi-agent systems. SPEAKER: Sunny Bakhda - Founding Engineer, HoneyHive 👉 Sign up for our "No BS" Newsletter to get the latest technical data & AI content: https://aicouncil.com/newsletter ABOUT AI COUNCIL: AI Council brings together the brightest minds in data to share industry knowledge, technical architectures and best practices in building cutting edge data & AI systems and tools. FIND US: Website: https://aicouncil.com/ LinkedIn: https://www.linkedin.com/company/aicouncilconf/ X: https://x.com/aicouncilconf
Read Video · Bản Ghi & Phân Tích
Tạo bản ghi & AI insight — Dùng thử miễn phí
Tài khoản miễn phí · không cần thẻ · 150 tín dụng khi đăng ký, đủ để mở khóa tập này
- 📄 Bản ghi đầy đủ có dấu thời gian
- ✨ Tóm tắt AI, từ khóa & bản đồ tư duy
- 💡 Điểm chính và trích dẫn nổi bật
Tập podcast và video sẵn sàng để đọc
Tập podcast

China's homegrown DUV lithography breakthrough and the debate over banning China's open-source AI
The AI Power Podcast
30 thg 7, 202656:33EN
50. Self-Checkout
The Economics of Everyday Things
22 thg 6, 202618:48EN
The gift and power of emotional courage | Susan David (re-release)
TED Talks Daily
29 thg 7, 202619:12EN
Por que o Talento Sozinho Não Vende Arte?
Art talks: Podcast do Paulo Varella
22 thg 2, 202614:17PT
Folge 7 - Kein Weg zurück: Flucht aus der DDR
Zeitreise DDR
14 thg 12, 202521:46DE
360-贝多芬是歌颂法国大革命的音乐家吗?
独树不成林
4 thg 8, 202638:09ZH-Hans
Video

QWEN just CRASHED the industry
Wes Roth
4 thg 8, 202615:21EN
Managed Agents - Don't Get Locked In
Sam Witteveen
13 thg 9, 202610:49EN
Agents Over Bubbles | Stratechery by Ben Thompson
Stratechery
26 thg 3, 202620:55EN
#39 Die Pest
99 mal Geschichte
8 thg 1, 20261:02:15DE
从「上瘾模型」到「专注力训练」,如何在被算法理解的世界里重新找回主动?| 英文访谈 S9E33
声动活泼
16 thg 10, 202550:48ZH-Hans