alerting-oncall

Configure alert rules, on-call rotations, and runbooks across monitoring stacks.

1|Updated Feb 5, 2026
One-click install
npx skills add https://github.com/allthingslinux/atl.services --skill alerting-oncall
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: alerting-oncall
Source: https://github.com/allthingslinux/atl.services/tree/main/.agents/skills/alerting-oncall
Command: npx skills add https://github.com/allthingslinux/atl.services --skill alerting-oncall

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill configures effective alerting and on-call management for production systems, helping teams reduce incident response time and alert fatigue.

Core Features & Use Cases

  • Configure alert routing and escalation policies across monitoring stacks (Prometheus, Alertmanager) and incident management platforms (PagerDuty, Opsgenie, Grafana OnCall).
  • Define on-call rotations, runbooks, and escalation guidelines to ensure timely response and clear ownership.
  • Provide templates and best practices for alert design, incident handling, and post-incident reviews to improve reliability.

Quick Start

Use the skill to outline your on-call rotation, create a basic alert routing rule, and draft a runbook for the most critical service.

Frequently Asked Questions about alerting-oncall

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I configure Prometheus alert routing and escalation policies in Alertmanager?

Configure Prometheus alert routing and escalation policies in Alertmanager by defining routing rules and grouping alerts to forward notifications to incident management platforms like PagerDuty, Opsgenie, or Grafana OnCall, reducing alert fatigue and improving incident response times.

What is the best way to set up an on-call rotation for production incident management?

The best way to set up on-call rotation for incident management is to define rotation schedules, escalation guidelines, and runbooks, ensuring clear ownership and timely response across integrated platforms like PagerDuty, Opsgenie, and Grafana OnCall.

Does this on-call workflow skill integrate with both PagerDuty and Grafana OnCall?

Yes, this on-call workflow skill integrates with PagerDuty, Opsgenie, and Grafana OnCall, allowing you to manage alert routing, configure escalation policies, and coordinate incident response across multiple monitoring stacks and incident management platforms.

How do I reduce alert fatigue when managing multiple monitoring stacks?

Reduce alert fatigue across monitoring stacks by applying best practices for alert design, configuring precise routing rules in Alertmanager, and establishing clear escalation policies to ensure only actionable alerts notify the on-call responder.

Can I generate runbooks and post-incident review templates for incident response?

Yes, you can generate runbooks and post-incident review templates for incident response, providing standardized best practices for alert handling and post-incident analysis to improve overall system reliability and on-call workflows.