Discussion about this post

User's avatar
Pradeep's avatar

Once AI workloads enter the mix, the old telemetry split becomes even more costly. An agent failure is rarely visible in one metric, one log, or one trace. You need the full chain: prompt version, tool calls, retrieval context, policy decisions, latency, and business outcome. Breaking that apart at ingest makes post-incident learning much weaker.

Anthony Joyes's avatar

I’ve followed Honeycomb as the “gold service” for as long as my aged brain allows me to recall. I’m stunned that we are still having these discussions :-)

No posts

Ready for more?