What problem does it solve?
This Skill helps you create new Claude Code skills, iteratively improve existing ones, and objectively measure whether a skill triggers correctly for the right user intents.
Core Features & Use Cases
- Skill drafting and iteration: Design a skill from scratch or refine an existing one using a structured workflow.
- Test-driven skill evaluation: Generate test prompts, run Claude-with-access-to-the-skill, and compare behavior against expected triggering.
- Quant + qualitative review loop: Draft quantitative evals when needed, then review results using the provided eval viewer and iterate until performance improves.
- Description optimization: Improve the SKILL.md description to increase triggering accuracy (reduce under-triggering and false triggers).
Quick Start
Use the skill creator to build a new skill by answering the questions about what the skill should do, when it should trigger, and what output format you want, then generate and run a small test set to start the eval/iterate loop.