incident-response

Generate incident response runbooks with severity definitions and mitigation steps.

1|Updated Mar 17, 2026
One-click install
npx skills add https://github.com/iceflower/agent-skills --skill incident-response-iceflower
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/iceflower/agent-skills/tree/main/incident-response
Command: npx skills add https://github.com/iceflower/agent-skills --skill incident-response-iceflower

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

Incident response workflow including severity classification, communication protocols, triage, mitigation strategies, runbook authoring, postmortem process, and on-call best practices. Covers MTTD, MTTA, MTTR metrics and SLO/SLI/SLA relationships. Use when handling production incidents, writing runbooks, or establishing incident response procedures.

Core Features & Use Cases

  • Severity classification and incident lifecycle templates
  • Runbook authoring and postmortem templates
  • On-call coordination, communication templates, and metrics dashboards
  • Alignment with blameless postmortems and SLOs

Quick Start

Describe a production incident scenario to generate a complete runbook, escalation plan, and post-incident processes.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I create an incident response runbook for production outages?

Generate a structured incident response runbook by describing a production outage scenario, which produces severity definitions, mitigation steps, communication templates, and post-incident postmortem checklists.

What is the best way to classify incident severity during on-call rotations?

Classify incident severity by evaluating impact against SLOs and SLIs, enabling on-call teams to trigger appropriate triage, mitigation, and escalation protocols during production outages.

How do I conduct a blameless postmortem after resolving an incident?

Conduct a blameless postmortem using generated templates to document MTTD, MTTA, and MTTR metrics, analyze root causes without assigning blame, and establish on-call best practices for future prevention.

Can I use this to generate communication templates for on-call coordination?

Yes, you can generate on-call coordination and communication templates that align with your incident lifecycle, ensuring structured information delivery during triage, mitigation, and post-incident reviews.

Does incident response planning work with existing SLO and SLA metrics?

Incident response planning aligns directly with SLO, SLI, and SLA metrics to define severity, measure MTTD and MTTR, and ensure mitigation strategies meet reliability objectives during production operations.