testing-skills-with-subagents

Test skills for pressure-resistance using TDD cycles and Serena metrics.

2|Updated Sep 30, 2025
One-click install
npx skills add https://github.com/krzemienski/shannon-framework --skill testing-skills-with-subagents-krzemienski
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: testing-skills-with-subagents
Source: https://github.com/krzemienski/shannon-framework/tree/main/skills/testing-skills-with-subagents
Command: npx skills add https://github.com/krzemienski/shannon-framework --skill testing-skills-with-subagents-krzemienski

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Provides a structured TDD workflow to verify process documentation and discipline with Serena metrics tracking, ensuring compliance scores improve with iterations.

Core Features & Use Cases

  • TDD cycle enforcement: RED-GREEN-REFACTOR for process testing.
  • Serena metrics: Track baseline failures, compliance, and rework.
  • Meta-testing support: Validate clarity and continuous improvement through meta-tests.

Quick Start

Run baseline scenarios, implement the skill to address failures, then re-run to achieve compliance_score > 0.85.

Frequently Asked Questions about testing-skills-with-subagents

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I test skills for robustness under pressure using TDD?

TDD-driven testing validates skill behavior through Red-Green-Refactor cycles with Serena metrics. Run baseline scenarios to identify failures, implement fixes, then re-run to achieve compliance_score > 0.85, ensuring bulletproof performance in high-pressure deployment environments.

What are Serena metrics and how do they measure compliance?

Serena metrics quantify compliance across scenarios involving time pressure, sunk cost, authority, exhaustion, and social pressures that may induce rationalizations. They track baseline failures, compliance scores, and loopholes to verify process discipline and continuous improvement through meta-testing.

How do I ensure my skill maintains compliance above 0.85 in subagent environments?

Log baseline failures, calculate compliance_score across pressure scenarios, track loopholes, and run meta-tests to validate prompt clarity. Iterate through Red-Green-Refactor cycles until the final score exceeds 0.85, ensuring consistent behavior across subagent deployment conditions.

What scenarios does this testing framework cover?

The framework tests skills against five pressure vectors: time pressure, sunk cost fallacy, authority influence, exhaustion, and social pressure. These scenarios simulate real deployment conditions where rationalization risks are highest, validating that documented processes hold under stress.

Can I use this testing method if I'm new to TDD?

The Skill enforces TDD discipline through structured Red-Green-Refactor cycles, making it accessible to TDD learners. Start with baseline scenarios, implement the skill to address failures, then validate compliance—the cycle itself teaches TDD fundamentals while building pressure-resistant skills.

What does meta-testing verify in this workflow?

Meta-testing validates that prompts and process documentation are clear and unambiguous. It catches rationalizations or loopholes in skill logic, ensuring the final compliance score reflects genuine robustness rather than incomplete coverage.