Skip to main content
Use the Monitoring views to understand how agents are used and how they perform over time. Track:
  • Message and run volume by agent, user, model, and time period.
  • Model and provider usage, token consumption, and cost.
  • Latency, tool activity, and failure rates.
  • Eval outcomes, judge signals, and user feedback.
Start with aggregate trends to identify a change in adoption, spend, or reliability. Then drill into the relevant run traces to find the concrete source of the change. This is the operational feedback loop for self-improvement: observe a pattern, diagnose the cause, make a governed change, and verify it with evals.