What problem does it solve?
Writing agent skills without real validation leads to rationalizations, loopholes, and instructions that fail under pressure.
Core Features & Use Cases
- Pressure-tested documentation: Convert TDD concepts into a “RED-GREEN-REFACTOR” loop for skill content so compliance is verified against baseline behavior.
- Loophole prevention: Identify exact rationalizations agents use when tempted to bypass rules, then encode explicit negations and red flags.
- Better discovery and selection: Enforce high-signal YAML frontmatter (name/description) and ensure the description specifies triggering symptoms rather than summarizing workflow.
- Use case: When you create or edit a discipline-enforcing skill (e.g., “must verify before proceeding”), this helps ensure it’s actually followed by other agents under time, authority, sunk-cost, and exhaustion pressure.
Quick Start
Use the writing-skills guide to draft your new skill, run a baseline “watch the skill fail” scenario, then update the skill until the same scenario passes.