incident-response

Coordinate triage, mitigation, and postmortem generation for production incidents.

Updated Feb 3, 2026
One-click install
npx skills add https://github.com/dhruvinrsoni/agentskills-garden --skill incident-response-dhruvinrsoni
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/dhruvinrsoni/agentskills-garden/tree/main/skills/90-maintenance/incident-response
Command: npx skills add https://github.com/dhruvinrsoni/agentskills-garden --skill incident-response-dhruvinrsoni

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill provides a structured, step-by-step process to effectively manage and resolve production incidents, minimizing downtime and capturing crucial learnings.

Core Features & Use Cases

  • Incident Triage: Classify severity, identify root causes, and isolate failing components.
  • Mitigation: Apply the fastest safe solutions like rollbacks or scaling.
  • Postmortem Generation: Create blameless postmortems for continuous improvement.
  • Use Case: When a critical service goes down, this Skill guides the team through immediate classification, diagnosis, and resolution, ensuring clear communication and a thorough post-incident review.

Quick Start

Use the incident-response skill to classify the current production issue as SEV2 and begin triage.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
What is the best way to manage production incident triage and severity classification?

Production incident triage requires a structured workflow to classify severity, identify root causes, and isolate failing components. This process ensures clear impact assessment and coordinates response efforts to minimize critical system downtime.

How do I resolve a critical system failure and apply immediate mitigation steps?

To resolve a critical system failure, you apply the fastest safe mitigation solutions like rollbacks or scaling. A structured incident response process guides the team through predefined decision criteria to swiftly restore service stability.

How do I generate a blameless postmortem after a production incident?

Generating a blameless postmortem after a production incident involves documenting the structured triage, mitigation, and resolution steps taken. This creates a thorough post-incident review for continuous improvement and future prevention.

Can I use this incident response workflow for SEV2 production support issues?

Yes, you can use this incident response workflow for SEV2 production support issues. It supports immediate response to critical system failures by guiding teams through severity classification, diagnosis, and mitigation.

When do I need a structured incident response process for troubleshooting?

You need a structured incident response process for troubleshooting when a critical service goes down. It provides predefined steps and decision criteria necessary for coordinating response efforts and minimizing system downtime.