What problem does it solve?
Developers and AI practitioners often lack a consistent, objective way to assess the quality of their agent skills. This leads to hidden weaknesses, poor emotional tone, and unreliable activation across sessions.
Core Features & Use Cases
- Type Classification – Automatically determines whether a skill is a reference, workflow, interactive, or principle skill.
- Comprehensive Scoring – Applies universal, type‑specific, and emotional‑tone rubrics to generate structural and affective scores.
- Actionable Recommendations – Provides concrete rewrite suggestions and prioritized improvement steps for any low‑scoring dimension.
- Test Scenario Generation – Produces realistic test cases to validate the evaluated skill after revisions.
Use case: a software engineer can run this evaluator on a newly authored workflow skill to receive a detailed report, fix tone issues, and obtain test scenarios before deployment.
Quick Start
Ask the skill evaluator to evaluate the skill located at './plugins/skill-authoring/skills/skill-evaluator'.