incident-response

Automate production incident response workflows from detection to postmortem.

Updated May 7, 2025
One-click install
npx skills add https://github.com/enchantednatures/.dotfiles --skill incident-response-enchantednatures
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/enchantednatures/.dotfiles/tree/main/.config/opencode/skills/incident-response
Command: npx skills add https://github.com/enchantednatures/.dotfiles --skill incident-response-enchantednatures

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Provide structured, repeatable incident response guidance to reduce mean time to recovery and improve blameless learning when production incidents occur.

Core Features & Use Cases

  • Incident response lifecycle templates (detection, triage, mitigation, recovery, and postmortem) with defined roles.
  • Playbooks, checklists, and communication templates to standardize coordination and status updates.
  • Reusable artifacts and runbooks that can be adapted for outages, performance degradations, security events, and third-party failures.

Quick Start

Activate the incident-response playbook immediately when a production incident is detected and follow the prescribed lifecycle from detection through postmortem.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I standardize incident response workflows to reduce mean time to recovery?

Standardize incident response by applying lifecycle templates from detection to postmortem. This provides defined roles, checklists, and communication plans to coordinate mitigation, preserve evidence, and reduce mean time to recovery.

What is the best way to structure a blameless postmortem after a production outage?

Structure a blameless postmortem using standardized reporting templates that enforce repeatable incident response cycles. This captures evidence and role responsibilities during recovery to improve auditability and continuous learning without assigning blame.

Can I use a single runbook for both security events and performance degradations?

Yes, reusable runbooks adapt across outages, performance degradations, security events, and third-party failures. They provide checklists and evidence collection templates to standardize actions regardless of the specific production incident type.

How do I enforce clear role responsibilities during on-call incident triage?

Enforce clear role responsibilities during on-call triage by activating a predefined incident response playbook. It codifies specific responder duties, communication templates, and triage checklists to standardize coordination and status updates.

Does incident response automation work without integrating external monitoring components?

Yes, incident response automation works independently by providing codified workflows and templates. It standardizes actions from detection through postmortem without requiring external components, focusing on runbooks and evidence collection for auditability.