What problem does it solve? Evaluating the maturity of an agent harness setup is subjective and inconsistent without a structured rubric. This Skill provides a repeatable scoring framework that grades harness readiness across nine dimensions and identifies the single next best improvement based on concrete evidence. ## Core Features & Use Cases - Nine-Dimension Scoring: Grades entry, context, state, feedback, evaluation, runtime observability, benchmark, cleanup, and plugin readiness on a 0-10 scale. - Evidence-Based Rules: Applies hard caps (e.g., missing startup path caps at 6, missing state or Judge caps at 7, unsafe auto-mutating plugins cap at 8) so scores reflect real gaps rather than optimism. - Structured Audit Output: Produces a markdown report with overall score, per-dimension scores, minimal harness gaps, capstone maturity gaps, and the next best improvement. - Use Case: After setting up an OpenCode harness configuration, run this audit to verify whether baseline readiness is met and which gap to close first before claiming capstone maturity. ## Quick Start Audit my current harness setup and score its maturity across all dimensions with evidence for each score.