incident-response

Detect downtime incidents, diagnose root causes, and coordinate fixes with budget guards.

1|1|Updated Apr 13, 2026
One-click install
npx skills add https://github.com/Cheggin/request-for-startups --skill incident-response-cheggin
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-response
Source: https://github.com/Cheggin/request-for-startups/tree/main/skills/incident-response
Command: npx skills add https://github.com/Cheggin/request-for-startups --skill incident-response-cheggin

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Automated incident management streamlines the full lifecycle from detection through diagnosis, fix, deployment, verification, and post-mortem, with escalation guarded by a budget.

Core Features & Use Cases

  • Detect incidents from uptime monitors or error-tracking signals and create traceable incident records.
  • Diagnose root causes using logs, traces, and metrics, then coordinate prioritized fixes and controlled deployments.
  • Verify recovery with health checks and structured post-mortems that include timelines, root cause, mitigations, and prevention.

Quick Start

Trigger an incident workflow by detecting a downtime event, diagnose the root cause, coordinate a fix and deployment, verify the outcome, and document a post-mortem within budget constraints.

Frequently Asked Questions about incident-response

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate incident management from detection to postmortem?

Automated incident management detects downtime or error spikes, diagnoses root causes, routes fixes, deploys changes, verifies recovery, and generates a postmortem with timelines and prevention recommendations within budget constraints.

What is the best way to diagnose root causes during an incident?

The best way to diagnose root causes during an incident is by analyzing logs, traces, and metrics to identify the failure source, then routing prioritized fixes to the appropriate owner for controlled deployment.

How do I create a tracked incident record from uptime monitor signals?

You create a tracked incident record by detecting downtime events or error spikes from uptime monitors and error-tracking signals, which automatically initializes a traceable incident workflow.

Can I enforce deployment budget guards during incident response?

Yes, you can enforce deployment budget guards during incident response to control escalation, ensuring that expedited changes and routing actions stay within predefined resource limits.

What should a structured postmortem include after incident recovery?

A structured postmortem should include the incident timeline, root cause analysis, applied mitigations, and specific prevention recommendations to document the event after health checks verify recovery.

How do I verify recovery after deploying an expedited incident fix?

You verify recovery after deploying an expedited fix by running health checks that confirm system stability, which then triggers the postmortem documentation phase within the incident management lifecycle.