agency-incident-response-commander

Coordinate production incident response with severity classification and blameless post-mortems.

Updated Jul 24, 2026
One-click install
npx skills add https://github.com/imMamdouhaboammar/kaku-chatgpt-harness --skill agency-incident-response-commander-immamdouhaboammar
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: agency-incident-response-commander
Source: https://github.com/imMamdouhaboammar/kaku-chatgpt-harness/tree/main/.agents/skills/engineering-incident-response-commander
Command: npx skills add https://github.com/imMamdouhaboammar/kaku-chatgpt-harness --skill agency-incident-response-commander-immamdouhaboammar

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill eliminates the chaos and uncertainty of production outages by providing a structured, blameless framework for incident detection, coordination, and post-mortem analysis.

Core Features & Use Cases

  • Structured Response: Implements a clear severity framework (SEV1-SEV4) and role-based coordination (IC, Comms, Tech Lead, Scribe).
  • Continuous Improvement: Facilitates blameless post-mortems using the 5 Whys to identify systemic root causes rather than individual errors.
  • Readiness & SLOs: Provides templates for runbooks, on-call rotation design, and SLO/SLI tracking to ensure long-term system reliability.

Quick Start

Use the agency-incident-response-commander skill to initiate a SEV2 incident response protocol for the checkout-api service.

Frequently Asked Questions about agency-incident-response-commander

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I coordinate a structured incident response for a production outage?

To coordinate structured incident response, you need a clear severity framework (SEV1-SEV4) and role-based coordination assigning Incident Commander, Comms, Tech Lead, and Scribe to manage real-time communication cadences and rapid resolution workflows.

What is a blameless post-mortem and when do I need to facilitate one?

A blameless post-mortem is a systemic root cause analysis process needed after any production incident, using techniques like the 5 Whys to identify systemic failures rather than individual errors, ensuring continuous improvement for distributed systems reliability.

How do I classify incident severity levels for high-availability distributed systems?

Incident severity classification for high-availability distributed systems uses a structured framework ranging from SEV1 to SEV4, satisfying requirements for rigorous SLO tracking and enabling rapid resolution workflows during production outages.

Can I use this incident response framework for on-call rotation design and SLO tracking?

Yes, this incident response framework provides templates for runbooks, on-call rotation design, and SLO/SLI tracking to ensure long-term system reliability for high-availability distributed systems requiring rigorous monitoring.

What is the best way to identify systemic root causes during a post-mortem?

The best way to identify systemic root causes during a post-mortem is facilitating blameless analysis using the 5 Whys technique, which reveals underlying systemic issues in distributed systems rather than attributing blame to individual errors.

Does this incident response protocol work for high-availability distributed systems requiring rigorous SLO tracking?

Yes, this incident response protocol is specifically designed for high-availability distributed systems requiring rigorous SLO tracking, providing structured coordination for incident detection, real-time communication cadences, and blameless post-mortem analysis.