What problem does it solve?
This Skill provides a comprehensive system for evaluating, tracking, and guiding the performance of all AI Agents within a project, ensuring quality, efficiency, and continuous improvement.
Core Features & Use Cases
- Multi-Agent Evaluation: Standardized scoring for all agents (programming, analysis, design, etc.).
- Performance Tracking: Monitors quality, token efficiency, and prediction accuracy over time.
- Feedback & Guidance: Provides tiered feedback from encouragement to critical warnings, with actionable recovery plans.
- Resource Management: Dynamically adjusts token budgets and task priorities based on agent performance.
- Use Case: Automatically assess an agent's code delivery, flag issues if it requires rework, and adjust its token budget for future tasks.
Quick Start
Evaluate the latest task completed by the Frontend Agent, noting its A-grade, 8/8 CHECKFIX pass, and 6.8k/8k token usage.