observability-methods

Map system performance issues to Four Golden Signals, RED, or USE methods.

2|Updated Mar 19, 2026
One-click install
npx skills add https://github.com/alex-voloshin-dev/ai-skills --skill observability-methods
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: observability-methods
Source: https://github.com/alex-voloshin-dev/ai-skills/tree/main/plugin/skills/observability-methods
Command: npx skills add https://github.com/alex-voloshin-dev/ai-skills --skill observability-methods

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill eliminates diagnostic ambiguity by providing a structured, industry-canonical vocabulary for investigating system performance, latency, and reliability issues.

Core Features & Use Cases

  • Methodology Mapping: Provides clear guidance on when to apply Four Golden Signals, RED, or USE methods based on the specific problem class.
  • Signal Correlation: Helps align SLI metrics, error budgets, and active alerts with the appropriate diagnostic framework.
  • Use Case: When investigating a production 5xx spike, use this skill to frame the analysis around the Four Golden Signals to ensure both error rates and resource saturation are evaluated simultaneously.

Quick Start

Apply the observability-methods skill to frame the current production incident investigation using the RED method for request-response analysis.

Frequently Asked Questions about observability-methods

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
What observability methodology should I use for diagnosing system latency and error spikes?

To diagnose system latency and error spikes, map the problem to industry-standard observability methodologies like the Four Golden Signals or RED method. This framework aligns specific problem classes such as request-response latency to the appropriate signal-based diagnostic method.

How do I standardize incident response analysis for production 5xx errors?

To standardize incident response for production 5xx errors, frame the investigation around the Four Golden Signals. This ensures both error rates and resource saturation are evaluated simultaneously, providing a structured, industry-canonical vocabulary for reliability issues.

When should I use the RED method versus the USE method for infrastructure monitoring?

The RED method is applied for request-response analysis, while the USE method is utilized for resource exhaustion. This framework provides clear guidance on choosing between them based on the specific problem class during production analysis.

How do I align SLI metrics and error budgets with active alerts during an incident?

To align SLI metrics and error budgets with active alerts, use a structured framework for signal correlation. This ensures your active alerts are evaluated within the appropriate diagnostic framework during system performance investigations.

Can I use this framework for resource exhaustion analysis in SRE workflows?

Yes, you can use this framework for resource exhaustion analysis in SRE workflows. It provides a structured methodology for diagnosing reliability issues by mapping specific problem classes like resource saturation to the appropriate signal-based diagnostic method.

What is the best way to investigate production system reliability without diagnostic ambiguity?

The best way to investigate production system reliability without diagnostic ambiguity is to apply a structured, industry-canonical vocabulary. This framework standardizes system analysis by mapping performance issues to signal-based diagnostic methods.