incident

Diagnose and mitigate incidents using a structured runbook with defined checkpoints.

Updated Apr 17, 2025
One-click install
npx skills add https://github.com/YazanKittaneh/blog.yazan.io --skill incident-yazankittaneh
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident
Source: https://github.com/YazanKittaneh/blog.yazan.io/tree/main/template/.claude/skills/incident
Command: npx skills add https://github.com/YazanKittaneh/blog.yazan.io --skill incident-yazankittaneh

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Incidents disrupt services and erode trust; this Skill provides a repeatable playbook to diagnose, mitigate, and document the resolution.

Core Features & Use Cases

  • Structured steps for declaring, diagnosing, mitigating, and communicating during incidents.
  • Templates for incident timelines and postmortems to enable blameless reviews.
  • Adaptable to web services, APIs, and databases across on-call scenarios.

Quick Start

Initiate an incident response runbook following this template for a newly detected outage.

Frequently Asked Questions about incident

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
What is a structured incident response runbook for managing outages?

An incident response runbook provides repeatable, step-by-step instructions to diagnose, mitigate, and communicate during service disruptions. It enforces diagnostic checks, mitigation steps, and timely updates for SREs handling outages, performance degradations, or security events.

How do I write a blameless postmortem after an outage?

To write a blameless postmortem, use a structured template to document the incident timeline and resolution. This approach enables objective reviews focusing on systemic issues rather than individual blame, ensuring continuous improvement across web apps, APIs, and databases.

How do I diagnose and mitigate performance degradation in APIs?

Diagnose and mitigate API performance degradation by following a structured runbook with defined checkpoints. This process applies diagnostic checks to identify root causes and executes mitigation steps to restore service stability for affected web applications.

Can I use this incident response playbook for database security events?

Yes, this incident response playbook is adaptable to database security events. It guides incident responders through declaring, diagnosing, and mitigating disruptions across on-call scenarios involving web services, APIs, and databases.

What is the best way to coordinate communications during an active incident?

The best way to coordinate incident communications is by following a structured runbook with defined checkpoints. This ensures timely, accurate updates are delivered to stakeholders throughout the diagnosis and mitigation phases of the outage.