What problem does it solve?
This Skill identifies and fixes logging and observability defects that make production systems difficult to diagnose, noisy to operate, expensive to monitor, or unsafe to expose.
Core Features & Use Cases
- End-to-End Observability Review: Trace critical journeys, failure paths, asynchronous work, telemetry pipelines, and operator workflows.
- Logging Quality and Safety Audit: Detect missing events, incorrect severity, excessive volume, unstable schemas, broken correlation, sensitive data exposure, and misleading messages.
- Production-Ready Remediation: Prioritize findings by operational impact, apply safe corrections, verify emitted telemetry, and document coverage gaps and residual risk.
- Use Case: Audit a distributed service before release to confirm that operators can reconstruct failures, distinguish retries from final outcomes, correlate work across queues and services, and investigate incidents without exposing secrets or overwhelming the telemetry backend.
Quick Start
Use the audit-logs skill to review logging and observability across the specified path, fix safe issues, and report findings with evidence, priorities, verification steps, and residual risks.