What problem does it solve?
This Skill provides a centralized, local dashboard to monitor the performance, cost, and reliability of a fleet of LLM agents, helping you identify bottlenecks and chain-of-thought failures without exposing sensitive data to external monitoring services.
Core Features & Use Cases
- Fleet Health Monitoring: Track call volume, latency (p50/p95), error rates, and costs across multiple agent archetypes.
- Trace Debugging: Visualize step-by-step tool call timelines to pinpoint exactly where an agent chain breaks.
- Human-in-the-loop: Flag specific traces or agents for investigation and maintain a local audit log of interventions.
- Use Case: Use this dashboard to review the performance of your support triage and booking assistant agents after a deployment to ensure error rates remain within acceptable thresholds.
Quick Start
Start the local observability dashboard by running the launcher script in the app directory and opening the provided local URL in your browser.