incident-response

Automate incident response with severity levels and runbooks for on-call teams.

6|Updated Dec 7, 2025
One-click install
npx skills add https://github.com/timequity/plugins --skill incident-response
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/timequity/plugins/tree/main/craft-coder/infra/incident-response
Command: npx skills add https://github.com/timequity/plugins --skill incident-response

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Structured incident response procedures with severity levels and on-call checklists to reduce MTTR.

Core Features & Use Cases

  • Severity & Flow: Clear response timings and an incident lifecycle.
  • On-Call Checklist: Acknowledge, assess, mitigate, resolve, and document.
  • Communication Templates: Standardized incident messages.

Quick Start

Reference the on-call checklist and runbooks to guide immediate investigation and resolution steps during incidents.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I reduce mean time to resolution (MTTR) during incidents?

Structured incident response with severity levels and standardized runbooks reduces MTTR by guiding on-call teams through a consistent workflow: Alert → Acknowledge → Assess → Mitigate → Resolve → Postmortem. Pre-built runbooks for common issues like High CPU and database slowness accelerate troubleshooting.

What should an on-call checklist include during incident handling?

An effective on-call checklist covers acknowledge, assess impact, mitigate the issue, resolve the root cause, and document findings. This ensures no critical step is missed and creates a record for postmortem analysis and team learning.

How do I set up escalation paths and severity levels for incidents?

Define severity levels tied to response timings and escalation paths that specify who to notify at each level. Clear severity criteria help on-call responders route issues appropriately and engage senior engineers only when needed, reducing false escalations.

Can I automate incident response for IT operations and SRE teams?

Yes. Incident response automation handles alert ingestion, impact assessment, runbook execution, and communication templates across IT operations and SRE workflows, allowing teams to respond consistently without manual coordination overhead.

What runbooks do I need for common infrastructure incidents?

Standard runbooks address High CPU, Out of Disk, and Database Slow scenarios. These provide step-by-step investigation and remediation procedures, reducing time spent searching for solutions during active incidents.

How do I standardize incident communication across my team?

Use communication templates within your incident response workflow to send consistent status updates, root cause summaries, and postmortem findings to stakeholders, improving transparency and reducing duplicate notifications.