What problem does it solve?
Guide production incident response for Kotlin and Spring services from alert through mitigation and follow-up, ensuring rapid containment and evidence preservation.
Core Features & Use Cases
- Reversible mitigation playbooks: rollback, feature flag adjustments, and load shedding designed for fast, safe containment.
- Evidence preservation and coordination: capture timelines, logs, metrics, and traces to enable root-cause analysis after stabilizing the system.
- Structured incident workflow: define incident severity, owners, and a clear sequence from alert to long-term remediation.
- Use Case: When an outage occurs after a deployment, apply this skill to stabilize service quickly, preserve evidence, and document next steps.
Quick Start
On incident onset, invoke Production Incident Responder to guide safe mitigation, preserve evidence, and structure the investigation.