incident-response

Triage and manage production incidents with classification, escalation, and communication templates.

1|Updated Mar 9, 2026
One-click install
npx skills add https://github.com/kiryteo/opencode-setup --skill incident-response-kiryteo
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/kiryteo/opencode-setup/tree/main/skills/incident-response
Command: npx skills add https://github.com/kiryteo/opencode-setup --skill incident-response-kiryteo

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill guides teams through structured incident response from detection to postmortem, reducing mean time to recover and improving reliability.

Core Features & Use Cases

  • Triage & classification: Quickly determine severity, scope, and incident commander roles.
  • Communication templates: Standardized status updates to stakeholders and customers.
  • Mitigation & resolution guidance: Step-by-step containment, repair, verification, and postmortem actions.
  • Use Case: During a critical outage, coordinate escalation, inform stakeholders, and document the incident timeline for the postmortem.

Quick Start

Describe the incident and I will guide you through triage, containment, and resolution steps.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I coordinate incident response during a production outage?

Incident response coordination involves triaging the outage, classifying severity, assigning an incident commander, and executing containment steps. This skill provides a repeatable workflow to manage escalations and standardize stakeholder communications.

What is the best way to conduct a blameless postmortem after an incident?

A blameless postmortem requires documenting the incident timeline, verifying resolution, and analyzing root causes without attributing fault. This skill guides you through capturing these details systematically to improve service reliability.

How do I triage production incidents and determine severity and scope?

Triage production incidents by assessing impact, determining severity, and defining the scope of degradation or outage. The skill helps quickly establish incident commander roles and escalation paths for scalable management.

Can I use this incident response workflow for service degradations and SLA breaches?

Yes, the incident response workflow is applicable to outages, service degradations, and SLA breaches across various platforms. It scales to handle different levels of severity and communication requirements.

How do I standardize stakeholder communication during an oncall incident?

Standardize stakeholder communication during an oncall incident by utilizing predefined status update templates. The skill includes communication templates to ensure consistent, accurate updates for stakeholders and customers.

What steps are involved in containing and verifying a production incident resolution?

Containing and verifying an incident involves executing step-by-step repair actions, confirming mitigation, and validating that the service is restored. The skill defines these containment and verification steps to safely close the incident.