From Spans to Trajectories: Observability for Long-Running Agents | HoneyHive
[2026 - DAY 2 - AI ENGINEERING] Agents have evolved. We've moved from orchestration frameworks — where agents operate through definite steps and turns you defined upfront — to harnesses, where the LLM uses skills and tools to chart its own trajectory. Modern agents run for hours or days, producing hundreds to thousands of steps in a single session. This calls for a fundamentally different methodology for monitoring and evaluating them in production. This talk shares what we've learned building observability infrastructure for agent harnesses at HoneyHive. We'll start with why the harness — not the model — has become the hardest engineering problem in production AI, and why traditional APM breaks down when traces are 10,000 spans deep and failures happen four tool calls deep. We'll walk through a live trajectory view to see what long-running agent traces actually look like at scale, and the specific challenges they create: context rot, semantic failure modes, and the needle-in-a-haystack problem of finding the moment that mattered. Then we'll dig into skills as the new unit of behavior and the dual role of clustering in agent development: unsupervised clustering for discovering emergent patterns and identifying where guardrails are needed, and supervised classifiers for production evaluation at scale. We'll close on what comes next — swarm observability for multi-agent systems. SPEAKER: Sunny Bakhda - Founding Engineer, HoneyHive 👉 Sign up for our "No BS" Newsletter to get the latest technical data & AI content: https://aicouncil.com/newsletter ABOUT AI COUNCIL: AI Council brings together the brightest minds in data to share industry knowledge, technical architectures and best practices in building cutting edge data & AI systems and tools. FIND US: Website: https://aicouncil.com/ LinkedIn: https://www.linkedin.com/company/aicouncilconf/ X: https://x.com/aicouncilconf
Read Video · 文字起こしと深掘り分析
文字起こしと AI インサイトを生成 — 無料体験
無料アカウント · カード不要 · 登録で150クレジット獲得、このエピソードのアンロックに十分
- 📄 タイムスタンプ付き全文文字起こし
- ✨ AI 要約・キーワード・マインドマップ
- 💡 重要ポイントと名言
すぐに読めるエピソードと動画
ポッドキャスト

Ep #82 Polar Explorers
Case by Case
2024年5月9日38:29EN
(Preview) Doom Debates Go Mainstream, AI Religion and the Economic Future, Several Vectors of the China Question
Sharp Tech with Ben Thompson
2026年9月18日33:19EN
How video games can level up the way you learn | Kris Alexander
TED Talks Daily
2026年9月7日14:25EN
#23 「人類好きなの?」
朝井リョウ・加藤千恵 信頼できない語り手
2026年7月10日56:14JA
COMO INTELIGÊNCIA ARTIFICIAL VAI QUEBRAR SEU NEGÓCIO SE VOCÊ NÃO OLHAR PARA ISSO | O Conselho 27
O Conselho
2025年7月17日1:11:11PT
vol.86 没事儿闯点小祸,反正闲着也是闲着
谐星聊天会
2026年7月8日1:23:40ZH-Hans
動画

Gemini Robotics – AI for the Physical World, with Keerthana and Ted of Google DeepMind
Cognitive Revolution "How AI Changes Everything"
2025年5月17日1:48:11EN
Jev: ChatGPT Co-Creator’s Answer to RLHF
AI Council
2026年6月19日36:39EN
Staphylococcus: Aureus, Epidermidis, Saprophyticus
Ninja Nerd
2021年10月21日1:01:18EN
Firma bez šéfov: Funguje to? - Money Talk 113 s Ferom Baníkom
Milan Dubec
2026年8月4日54:10SK
2026/08/24(一) 輝達伺服器傳漲價15%:AI成本暴增,成本誰吸收?
財女珍妮
2026年8月24日30:06ZH-Hant