slo-implementation

Define and implement SLIs and SLOs with error budgets and alerting.

Updated Apr 19, 2026
One-click install
npx skills add https://github.com/ArogyaReddy/https-github.com-wshobson-agents --skill slo-implementation-arogyareddy
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: slo-implementation
Source: https://github.com/ArogyaReddy/https-github.com-wshobson-agents/tree/main/plugins/observability-monitoring/skills/slo-implementation
Command: npx skills add https://github.com/ArogyaReddy/https-github.com-wshobson-agents --skill slo-implementation-arogyareddy

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

Defines and implements SLIs (Service Level Indicators) and SLOs (Service Level Objectives) with error budgets and alerting to drive reliability.

Core Features & Use Cases

  • Framework for designing SLIs (availability, latency, etc.) and SLOs with clear error budgets.
  • Includes Prometheus recording rules and alerting patterns to enforce reliability targets.
  • Use Case: establish service reliability targets, measure user-visible performance, and automate incident responses.

Quick Start

Configure your service with the provided SLO templates and apply the Prometheus-based rules to start monitoring reliability.

Frequently Asked Questions about slo-implementation

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I define SLIs and SLOs with error budgets for my services?

Defining SLIs and SLOs involves establishing reliability targets and measuring user-perceived performance. This framework provides templates to define indicators like availability and latency, paired with clear error budgets to drive alerting and incident responses.

How do I set up Prometheus alerting rules based on SLO error budgets?

You can set up Prometheus alerting rules based on SLO error budgets using the provided recording rules and alerting patterns. These rules enforce reliability targets and automate incident responses when error budgets are exhausted.

What is the best way to measure user-perceived performance and service reliability?

The best way to measure user-perceived performance is by implementing SLIs and SLOs across your services. This approach applies structured reliability targets and error budgets to track user-visible performance and drive incident responses.

Can I use Grafana dashboards to visualize SLO error budgets and service reliability?

Yes, you can use Grafana dashboards to visualize SLO error budgets and service reliability. The framework supports Grafana dashboards alongside Prometheus recording rules to monitor and review your established reliability targets.

Do I need Prometheus to implement SLO alerting and recording rules?

Prometheus is required to implement the provided SLO alerting and recording rules. The templates are designed specifically for Prometheus-based monitoring to enforce reliability targets and automate incident responses.

How do I conduct a structured SLO review process for service reliability targets?

You conduct a structured SLO review process using the included documentation references. This helps evaluate your SLIs and SLOs, measure user-visible performance, and adjust error budgets and alerting patterns as needed.