What problem does it solve? Production services break in unanticipated ways, and teams without a documented incident response process improvise under pressure, leading to slow resolution, poor communication, and repeated failures. This Skill produces a complete incident response framework so teams know exactly how to declare, coordinate, resolve, and learn from incidents. ## Core Features & Use Cases - Severity Matrix & Declaration Process: Generates SEV-1/2/3 definitions with response SLAs, escalation rules, and Slack-based incident declaration commands including the Incident Commander role. - Blameless Postmortem Template: Produces a Google SRE-standard postmortem covering summary, timeline, root cause analysis, impact metrics, and corrective actions with owners and due dates. - MTTD/MTTR Tracking & Action-Item Closure: Creates an incident log, feeds DORA change failure rate and MTTR metrics via delivery-record scripts, and tracks open postmortem action items as a reliability metric. - Resilience Verification: Adds DiRT-style game day drills and Wheel of Misfortune role-plays to test runbooks and on-call readiness before real outages. - Use Case: After setting up SLOs and runbooks for a new production API, run this Skill to generate the incident response process, postmortem template, and tracking log, then validate everything with the incident-response-reviewer agent. ## Quick Start Set up a complete incident response process for my production service, including severity levels, a postmortem template, and MTTR tracking.