systematic-debugging

Diagnose bugs through a four-phase root cause investigation process before proposing fixes.

Updated Apr 29, 2026
One-click install
npx skills add https://github.com/fred-meng/harness --skill systematic-debugging-fred-meng
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: systematic-debugging
Source: https://github.com/fred-meng/harness/tree/main/.github/skills/systematic-debugging
Command: npx skills add https://github.com/fred-meng/harness --skill systematic-debugging-fred-meng

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve? Random fixes and quick patches waste time, mask underlying issues, and introduce new bugs. This Skill enforces a disciplined debugging methodology that finds the actual root cause before any fix is attempted, even under time pressure or social pressure to apply a quick workaround. ## Core Features & Use Cases - Four-Phase Process: Root cause investigation, pattern analysis, hypothesis testing, and implementation, with mandatory completion of each phase before proceeding. - Supporting Techniques: Includes root-cause-tracing for backward call-stack analysis, defense-in-depth validation at multiple layers, and condition-based waiting to replace flaky arbitrary timeouts in tests. - Pressure Resistance: Explicit anti-patterns, red flags, and rationalization tables that stop shortcut fixes during emergencies, plus a rule to question the architecture after three failed fix attempts. - Use Case: When a production test fails intermittently, use this Skill to trace the failure back through the call chain, form a single testable hypothesis, create a failing test case, and fix the source rather than adding another sleep timeout. ## Quick Start Use the systematic-debugging skill to investigate this failing test and find its root cause before suggesting any fix.

Frequently Asked Questions about systematic-debugging

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I debug a test failure systematically instead of guessing?

Follow the four phases: investigate the root cause by reading errors and reproducing consistently, analyze patterns against working examples, form and test a single hypothesis, then implement one fix with a failing test case. Never propose fixes before completing the investigation phase.

How to find the root cause of a bug deep in the call stack?

Use root cause tracing: observe the symptom, find the immediate cause, then trace backward through each caller until you find where the bad value originated. Add stack trace instrumentation with console.error before the failing operation if manual tracing is not possible.

How do I fix flaky tests that use setTimeout or sleep?

Replace arbitrary timeouts with condition-based waiting that polls for the actual condition you care about, such as an event appearing or state changing. Poll every 10ms with a clear timeout error, and only use fixed delays when testing actual timing behavior.

What should I do when my first fix does not work?

Stop and return to Phase 1 to re-analyze with the new information rather than stacking more fixes. If three or more fixes have failed, treat it as an architectural problem and discuss the fundamental design with your team before attempting another fix.

When is it acceptable to skip root cause investigation for a quick fix?

Never, according to this methodology. Even under production emergencies or time pressure, systematic investigation is faster than guess-and-check thrashing, and symptom fixes are treated as failures that mask the real issue.

How do I find which test is polluting shared state or creating files?

Use the included find-polluter.sh bisection script, which runs test files one by one and checks whether the unwanted file or directory appears after each run. It stops at the first polluting test and reports the file for investigation.