What problem does it solve?
This Skill solves the common issue of monitoring dashboards that are cluttered with irrelevant "vanity" metrics and fail to help operators quickly answer critical questions about system health, bottlenecks, recent changes, and required corrective actions.
Core Features & Use Cases
- Operator-question-first workflow: Structures dashboards around core operational categories (health, latency, throughput, saturation, service-specific risks) instead of random metric inclusion.
- Multi-platform guidance: Provides step-by-step instructions for building dashboards on Grafana, SigNoz, and similar monitoring platforms, including schema inspection and panel best practices.
- Pre-built templates and checklists: Includes example panel sets for common services (Elasticsearch, Kafka, API gateways) and a quality checklist to ensure dashboards are usable for operations.
- Use Case: If you need a Kafka monitoring dashboard, this Skill guides you to include only relevant panels like under-replicated partitions, consumer lag, and disk pressure, rather than every available broker metric.
Quick Start
Use the dashboard-builder skill to create a Grafana dashboard for your Elasticsearch cluster that answers questions about cluster health, shard allocation, and JVM pressure.