What problem does it solve?
This Skill helps you prevent alert noise and missed responses by providing a practical blueprint for Grafana Alerting and Incident Response Management (IRM), including SLO burn-rate alerts and end-to-end notification routing.
Core Features & Use Cases
- Grafana-managed alerting provisioning: Define alert groups, rules, expressions, and lifecycle states (including NoData/Error handling) in YAML so alerting can be deployed and reviewed like code.
- Data source-managed rule examples: Cover Prometheus/Mimir ruler-style alert rules (including recording rules) and Loki LogQL-based alerting for log-driven incidents.
- Notification routing for on-call: Set up contact points (Slack/PagerDuty/email/webhook), build notification policies, and use silences/mutes to control who gets paged and when.
- SLO configuration with burn-rate alerts: Generate SLO-related recording rules and burn-rate alerting logic for availability and error budgets.
- IRM on-call workflows: Outline on-call schedules, escalation chains, incident collaboration, and integration sources to connect alerts to operational response.
Quick Start
Provision Grafana alert rules, contact points, and notification policies by generating the YAML templates for alerting/routing, then apply them via Grafana provisioning APIs so incidents and SLO burn-rate alerts route to the correct on-call channels.