incident-mode

Identify incident scope, impact, and environment for safe remediation planning.

Updated Dec 25, 2025
One-click install
npx skills add https://github.com/jibaxZZZ/codex-root-configuration --skill incident-mode
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-mode
Source: https://github.com/jibaxZZZ/codex-root-configuration/tree/main/.codex/skills/incident-mode
Command: npx skills add https://github.com/jibaxZZZ/codex-root-configuration --skill incident-mode

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Triaging production incidents is time-consuming and error-prone; this Skill provides a safe, structured approach with rollback guidance to minimize blast radius.

Core Features & Use Cases

  • Scoped incident assessment and impact analysis
  • Safe remediation planning with explicit rollback steps
  • On-call documentation and post-incident follow-ups

Quick Start

Stabilize a failing service by quickly assessing scope, logging findings, and drafting a rollback plan.

Frequently Asked Questions about incident-mode

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I triage a production incident safely?

To triage a production incident safely, identify the incident scope, impact, and environment to streamline remediation, while specifying procedural steps and rollback planning with safety guardrails to minimize blast radius.

What is the best way to plan a rollback during an on-call outage?

The best way to plan a rollback during an on-call outage is to follow a structured remediation plan with explicit rollback steps and safety guardrails, ensuring you assess the incident scope and impact before executing the rollback.

How do I assess the impact and scope of a production outage?

Assess the impact and scope of a production outage by identifying the affected services, environments, and production ecosystems to streamline safe remediation and minimize the blast radius of the incident.

Can I use this approach for on-call outages across different environments?

Yes, this approach is applicable to on-call outages across services, environments, and production ecosystems, providing a structured method to assess incident scope and plan safe remediation regardless of the specific environment.

What should I include in post-incident follow-ups after a production outage?

Post-incident follow-ups after a production outage should include on-call documentation and procedural steps that capture the incident scope, impact analysis, and rollback planning to ensure future safety and streamlined remediation.

Why does incident triage need explicit safety guardrails?

Incident triage needs explicit safety guardrails because triaging production incidents is error-prone, and safety guardrails ensure structured remediation with rollback guidance to minimize blast radius and prevent further impact.