incident-response

Triage production incidents, draft status updates, and write blameless postmortems.

2|1|Updated May 17, 2026
One-click install
npx skills add https://github.com/rakibulism/agent-skills-os --skill incident-response-rakibulism
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/rakibulism/agent-skills-os/tree/main/skills/engineering-incident-response
Command: npx skills add https://github.com/rakibulism/agent-skills-os --skill incident-response-rakibulism

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill helps teams handle production incidents from first alert to resolution by structuring triage, communication, mitigation, and postmortem work in one clear workflow.

Core Features & Use Cases

  • Incident Triage: Assess severity, identify impacted systems and users, and assign roles so the team can respond quickly.
  • Live Communication: Draft concise internal and customer-facing status updates with current impact, actions taken, and next steps.
  • Postmortem Writing: Reconstruct timelines, analyze root cause, and capture blameless action items after the incident is resolved.
  • Use Case: When a monitoring alert indicates an outage, use this Skill to classify severity, produce the first status update, and later generate the postmortem document.

Quick Start

Use the incident response skill to triage this outage, draft the current status update, and prepare a blameless postmortem outline.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I triage a production outage and classify severity?

Triage a production outage by assessing severity, identifying impacted systems and users, and assigning incident response roles so the team can respond quickly. Structured triage ensures rapid mitigation and clear ownership during degraded service.

What is a blameless postmortem and how do I write one?

A blameless postmortem reconstructs the incident timeline, analyzes root cause, and captures action items without assigning blame. Generate one after resolution by reviewing factual updates and documenting preventive measures across operational teams.

How do I draft status updates during an active incident?

Draft status updates during an active incident by summarizing current impact, actions taken, and next steps. Concise internal and customer-facing updates keep stakeholders informed throughout the outage investigation and resolution process.

What's the best way to coordinate incident response across operational teams?

Coordinate incident response across operational teams by applying structured severity classification, timeline tracking, and factual update drafting. This workflow manages alert investigations and live communication from first alert to resolution review.

When do I need a structured incident response process?

You need a structured incident response process when monitoring alerts indicate outages, degraded service, or require alert investigations. It manages triage, communication, mitigation, and postmortem documentation across operational teams from first alert to resolution.

Can I use this incident response workflow for alert investigations and resolution reviews?

Yes, this incident response workflow applies to alert investigations and resolution reviews. It structures severity classification, timeline tracking, and postmortem documentation to manage outages and degraded service across operational teams.