← All field notes

Topic

Agent Evals & Observability

If you cannot see the tool calls, you cannot trust the agent. This hub is about traces as ground truth, evals that match on-call reality, local models scored against production LangSmith runs, and observability that scales past hero engineers — including how OpenTelemetry fits an AgentOps stack.