Agent observability needs action-level traces

Knowing an agent ran is useless. Knowing what it decided, called and received is what lets you investigate.

Most deployments log the request and the response, which is the equivalent of recording that a conversation happened. The intermediate steps, what was retrieved, which tool was invoked with what arguments, what came back, are where both the failure and the manipulation are visible.

More on AI agents