incident-triage-automator

Aggregate alerts from monitoring systems and compute incident severity.

2|1|Updated Mar 11, 2026
One-click install
npx skills add https://github.com/lloydchang/agentic-reconciliation-engine --skill incident-triage-automator
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: incident-triage-automator
Source: https://github.com/lloydchang/agentic-reconciliation-engine/tree/main/core/ai/skills/incident-triage-automator
Command: npx skills add https://github.com/lloydchang/agentic-reconciliation-engine --skill incident-triage-automator

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Automates the triage and management of operational incidents by aggregating alerts, assessing severity, creating incidents, notifying stakeholders, and tracking resolution progress to reduce MTTR and ensure consistent incident handling.

Core Features & Use Cases

  • Alert aggregation from Prometheus, Datadog, PagerDuty, Grafana, and other monitoring sources.
  • Severity assessment and automated incident creation with runbook recommendations.
  • Stakeholder notification across Slack, Email, Microsoft Teams, and PagerDuty, plus status tracking and post-mortem generation.
  • Runbook-driven remediation guidance and escalation policy enforcement for P0–P3 incidents.

Quick Start

Describe an incident scenario to automatically triage, assess severity, and escalate to the appropriate responders.

Frequently Asked Questions about incident-triage-automator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate incident triage for alerts from multiple monitoring systems?

You can automate incident triage by aggregating alerts from Prometheus, Datadog, PagerDuty, and Grafana, then computing severity to create incidents. This Skill assesses severity, recommends runbooks, and tracks resolution progress to reduce MTTR and ensure consistent incident handling.

What is alert aggregation and severity assessment in incident management?

Alert aggregation and severity assessment in incident management consolidates alerts from multiple monitoring sources and computes incident severity. This process supports automated incident creation, stakeholder notification, and runbook recommendations for consistent operational response.

Does this incident triage automator work with Prometheus, Datadog, and PagerDuty?

Yes, this incident triage automator works with Prometheus, Datadog, PagerDuty, and Grafana. It aggregates alerts from these monitoring systems to compute severity, create incidents, notify stakeholders, and enforce configurable escalation policies for P0 through P3 incidents.

How do I set up escalation policies for P0 to P3 incidents?

Setting up escalation policies for P0 to P3 incidents involves defining severity-based rules for stakeholder notification and runbook-driven remediation. This Skill enforces configurable escalation policies across Slack, Email, Microsoft Teams, and PagerDuty to track status and guide response.

What's the best way to notify stakeholders during a production IT incident?

The best way to notify stakeholders during a production IT incident is automated multi-channel notification across Slack, Email, Microsoft Teams, and PagerDuty. This Skill triggers stakeholder notifications based on computed severity and tracks resolution status to ensure consistent incident handling.

Can I get runbook recommendations for automated incident response?

Yes, you can get runbook recommendations for automated incident response. This Skill provides runbook-driven remediation guidance during incident triage, applying to production IT operations and alert-driven response workflows to reduce mean time to resolution.