incident-response

Triage and manage production incidents from detection to resolution.

Updated Mar 15, 2026
One-click install
npx skills add https://github.com/lilbom32/ketnoitrithuc --skill incident-response-lilbom32
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/lilbom32/ketnoitrithuc/tree/main/.claude/skills/engineering/1.1.0/skills/incident-response
Command: npx skills add https://github.com/lilbom32/ketnoitrithuc --skill incident-response-lilbom32

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This tool helps engineering teams triage and manage production incidents efficiently, reducing MTTR and coordinating responders during outages.

Core Features & Use Cases

  • Severity classification and incident commander assignment to ensure clear ownership.
  • Real-time communication templates and status updates to keep stakeholders informed.
  • Mitigation, containment, resolution workflows, and postmortem action tracking.

Quick Start

Trigger the incident-response workflow by describing the incident and its severity to initiate triage and coordination.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I triage a Sev1 production outage and assign an incident commander?

Real-time incident communication templates provide structured status updates to keep stakeholders informed during outages. They standardize the frequency and format of broadcasts, ensuring consistent messaging across teams while the incident commander coordinates mitigation efforts.

What is the best way to manage a postmortem and track action items after an incident?

Yes, incident response workflows support Sev1 through Sev4 outages. Rapid severity classification determines the required escalation, assigns appropriate roles like the incident commander, and tailors the mitigation, containment, and resolution workflows based on the incident's impact.

How does incident response coordinate multiple teams during an on-call page?

Trigger the incident response workflow by describing the incident and its severity. This initiates rapid triage, assigns an incident commander for clear ownership, and launches standardized mitigation, communication, and resolution workflows to coordinate stakeholders end-to-end.

Does this incident response workflow handle containment and verification across different services?

Post-incident reviews are conducted using a standardized postmortem process that documents the incident timeline and root cause. This structured review tracks action items and verifies resolutions to ensure continuous improvement and prevent future outages across services.

Can I use this for on-call pages across multiple systems and services?

Yes, you can use this for on-call pages across multiple systems and services. The incident response workflow manages triage, communication, and mitigation end-to-end, coordinating stakeholders regardless of the specific underlying service architecture.