benchmarking-dashboard

Monitor and visualize AI agent performance metrics with persistent dashboards.

Updated Mar 12, 2026
One-click install
npx skills add https://github.com/ryasrk/AgentBrokeTheMatrix-CopilotVersion --skill benchmarking-dashboard
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: benchmarking-dashboard
Source: https://github.com/ryasrk/AgentBrokeTheMatrix-CopilotVersion/tree/main/.github/skills/benchmarking-dashboard
Command: npx skills add https://github.com/ryasrk/AgentBrokeTheMatrix-CopilotVersion --skill benchmarking-dashboard

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill provides a centralized, persistent dashboard to track and analyze the performance of AI agents, helping identify utilization patterns, areas for improvement, and overall system health.

Core Features & Use Cases

  • Metric Tracking: Logs token usage, dispatch counts, routing accuracy, and session scores.
  • Agent Utilization Analysis: Identifies high-value, underutilized, or unused agents.
  • Trend Visualization: Tracks performance trends over time for routing accuracy and context efficiency.
  • Use Case: After a week of development using various agents, review the dashboard to see which agents were most effective, if the planner is routing tasks correctly, and if context is being used efficiently.

Quick Start

Review the agent benchmarks dashboard to understand current system performance.

Frequently Asked Questions about benchmarking-dashboard

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I track AI agent performance metrics like token usage and routing accuracy?

To track AI agent performance metrics, you need a centralized dashboard that logs token usage, dispatch counts, routing accuracy, and session scores. This allows you to identify utilization patterns, find underperforming agents, and monitor overall system health.

What is the best way to analyze agent utilization and identify redundant agents?

Analyzing agent utilization involves reviewing dispatch counts and session scores to identify high-value, underutilized, or unused agents. A performance dashboard helps pinpoint redundant agents by visualizing their activity levels and effectiveness over time.

How do I visualize AI agent performance trends over time for context efficiency?

To visualize AI agent performance trends, use a dashboard that tracks routing accuracy and context efficiency over historical sessions. This requires persistent storage for historical data to facilitate analysis of system health trends post-session.

Does tracking agent performance metrics require persistent storage for historical data?

Yes, tracking agent performance metrics requires persistent storage for historical data to log sessions accurately and enable automated updates post-session. This storage ensures you can review system health trends and routing accuracy over time.

Can I review which AI agents were most effective after a week of development?

Yes, you can review which AI agents were most effective after development by checking a benchmarks dashboard. It evaluates if the planner routes tasks correctly, tracks token usage, and shows whether context is being used efficiently.

Why does monitoring system health require automated updates post-session?

Monitoring system health requires automated updates post-session to ensure the dashboard accurately reflects the latest token usage and routing accuracy. This automated process maintains current agent utilization data for reliable trend visualization.