harness-metrics-audit

Define real success metrics and regression checks for a Claude Code harness.

Updated Jun 30, 2026
One-click install
npx skills add https://github.com/Festo-Wampamba/Claude-Features --skill harness-metrics-audit
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: harness-metrics-audit
Source: https://github.com/Festo-Wampamba/Claude-Features/tree/main/skills/harness-metrics-audit
Command: npx skills add https://github.com/Festo-Wampamba/Claude-Features --skill harness-metrics-audit

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill helps you figure out whether your Claude Code harness is actually getting better in the way you care about, instead of only improving easy-to-count proxy metrics like skill count, automation coverage, or token savings.

Core Features & Use Cases

  • Clarifies the real target: Interviews the user to define what better means in plain language, such as faster delivery, fewer corrections, or higher trust.
  • Separates signal from proxies: Compares visible metrics against the outcome they are supposed to represent and exposes when they drift apart.
  • Designs regression checks: Proposes concrete evals for real workflows, friction counters, trust checks, and time or cost measurements.
  • Use Case: You changed your hooks and skills and want to know whether the setup is genuinely helping, not just looking more automated.

Quick Start

Ask the skill to interview your goals, identify proxy metrics that may be misleading, and produce a small regression suite for your Claude Code harness.

Frequently Asked Questions about harness-metrics-audit

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I measure if my Claude Code harness is actually improving?

To measure if your Claude Code harness is improving, you must define real success targets and compare them against visible proxy metrics. This skill interviews your goals and exposes when easy-to-count metrics drift from the outcomes you care about.

What is proxy-versus-target analysis in automation metrics?

Proxy-versus-target analysis compares visible metrics like skill count or token savings against the real outcome they represent. It detects when your automation optimizes the wrong outcome by exposing the gap between the proxy and the actual goal.

How do I set up regression checks for Claude Code hooks and skills?

You set up regression checks by designing concrete evaluations for real workflows, friction counters, and trust checks. This skill produces a small regression suite to test whether configuration changes genuinely help.

Why do my automation metrics improve but user satisfaction drops?

Automation metrics improve while user satisfaction drops because proxy metrics drift from real outcomes. When you optimize easy-to-count metrics like automation coverage, you may be optimizing the wrong outcome without realizing it.

Can I audit my Claude Code configuration changes without external dependencies?

Yes, you can audit Claude Code configuration changes without external dependencies. This skill requires no dependencies to interview your goals, identify misleading proxies, and produce a review cadence to detect drift.

When do I need a harness metrics audit for my Claude Code setup?

You need a harness metrics audit when automation or speed may be optimizing the wrong outcome. It applies to harness audits, configuration changes, hooks, skills, and regressions to detect drift before users feel it.