Topic
Agent Evals & Observability
If you cannot see the tool calls, you cannot trust the agent. This hub is about traces as ground truth, evals that match on-call reality, local models scored against production LangSmith runs, and observability that scales past hero engineers — including how OpenTelemetry fits an AgentOps stack.