Observability Designer

Design SLI/SLO frameworks, alert optimization, and dashboards for production systems.

Updated Mar 7, 2026
One-click install
npx skills add https://github.com/tapanshah/Claude-Skills --skill observability-designer-tapanshah
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: Observability Designer
Source: https://github.com/tapanshah/Claude-Skills/tree/main/engineering/observability-designer
Command: npx skills add https://github.com/tapanshah/Claude-Skills --skill observability-designer-tapanshah

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) and assets (resource) components.

What problem does it solve?

This Skill addresses the complexity of designing and implementing effective observability strategies, ensuring systems are monitored comprehensively and alerts are actionable.

Core Features & Use Cases

  • SLI/SLO Frameworks: Design Service Level Indicators and Objectives tailored to service criticality.
  • Alert Optimization: Analyze and refine alert configurations to reduce noise and improve signal.
  • Dashboard Generation: Create role-based dashboards for SREs, developers, and executives.
  • Use Case: A platform engineering team can use this Skill to automatically generate a complete observability strategy, including SLIs, SLOs, alert rules, and Grafana dashboards, for a new microservice, significantly reducing manual setup time and ensuring best practices are followed.

Quick Start

Use the Observability Designer skill to generate an SLO framework for a new critical API service.

Frequently Asked Questions about Observability Designer

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I design an SLO framework for a new microservice?

To design an SLO framework for a microservice, this Skill analyzes service characteristics to generate tailored Service Level Indicators and Objectives based on service criticality. It applies best practices for metrics, logs, and traces.

What is the best way to optimize alerting rules to reduce noise?

The best way to optimize alerting rules is to analyze and refine alert configurations to reduce noise and improve signal. This Skill refines alert rules to ensure alerts are actionable for production systems.

Can I generate Grafana dashboards for different roles automatically?

Yes, you can automatically generate Grafana dashboards for different roles. This Skill creates role-based dashboards tailored for SREs, developers, and executives as part of a complete observability strategy.

Does this observability strategy design work with Prometheus monitoring stacks?

Yes, this observability strategy design works with Prometheus. The Skill integrates with common monitoring stacks like Prometheus and Grafana to automate the creation of monitoring solutions.

How do I create monitoring solutions that cover metrics, logs, and traces?

To create monitoring solutions covering metrics, logs, and traces, this Skill applies observability best practices by analyzing your service characteristics. It ensures systems are monitored comprehensively across all telemetry signals.