What problem does it solve? It is hard to know whether installed agent skills actually help. This Skill scores recent local agent conversations against efficiency and code-quality rubrics, identifies which skills fired and which failed, and drafts concrete skill edits with a shareable report. ## Core Features & Use Cases - Conversation Scoring: Samples local Claude Code, Codex, and Warp sessions and grades each transcript on efficiency and code quality using defined rubrics. - Skill Coverage Analysis: Detects which installed skills were actually used and computes a weighted overall grade. - Drafted Skill Edits: Produces improved SKILL.md versions with unified diffs, traced back to failed conversations, without touching real skill files. - Shareable Report: Renders a self-contained HTML report with a letter grade, findings, suggestions, and a PNG export button. - Use Case: After a month of agent-assisted work, run the skill to learn that your test-runner skill never triggered and receive a rewritten trigger description that would have fired in three failed sessions. ## Quick Start Ask the agent to grade my recent agent conversations and tell me which of my installed skills are actually working.