AI agents can be fast, available, and still completely wrong. Standard application monitoring can show latency and errors, but not whether an answer was useful, a tool call behaved correctly, or a prompt change introduced a regression. This talk shows how to instrument, trace, and evaluate production AI agents using OpenTelemetry and AI Observability.
This talk has been presented at AI Coding Summit NYC, check out the latest edition of this Tech Conference.


















