What problem does it solve?
This Skill helps you triage a live production incident by assessing severity, containing blast radius, and coordinating a clean handoff without prematurely chasing root cause.
Core Features & Use Cases
- Incident severity classification to quickly decide who to page and how aggressively to stop other work.
- Blast-radius assessment and mitigation-first actions to restore service using the smallest effective intervention (rollback, scaling, feature flags, rate limiting, routing changes).
- Operational communication and handoff discipline with a consistent incident timeline, regular updates, and a scheduled post-mortem.
Use case: A new deploy causes elevated errors and latency; you need to decide Sev1 vs Sev2, mitigate immediately (rollback/traffic drain/flag), keep stakeholders updated, and transfer context for resolution.
Quick Start
Ask your AI assistant to run triage-incident for the current on-call alert, producing a timeline, current state, known facts, unclear items, and next mitigation actions.