What problem does it solve? It is hard to know whether installed agent skills actually help. This Skill grades your agent setup from real local conversation history, scoring efficiency and code quality, then proposes concrete, evidence-backed skill edits. ## Core Features & Use Cases - Conversation Scoring: Samples recent Claude Code, Codex, Pi, and Warp sessions and scores them against efficiency and code-quality rubrics. - Skill Improvement Drafts: Writes proposed SKILL.md edits with unified diffs, each traced to a failed conversation as evidence. - Shareable Report: Renders a self-contained HTML report with a letter grade, score bars, findings, and a share-as-PNG button. - Use Case: After a month of coding with an agent, run the skill to learn that your debugging skill never triggers, and receive a rewritten trigger description plus a diff you can apply. ## Quick Start Ask the agent to use the skill-doctor skill to grade my recent agent conversations and suggest improvements to my installed skills.