What problem does it solve?
This Skill helps engineers operate and troubleshoot OSDC observability when monitoring data, system logs, Kubernetes events, credentials, or Grafana Cloud queries are difficult to configure or diagnose.
Core Features & Use Cases
- Metrics Operations: Configure and troubleshoot Alloy, kube-prometheus-stack, ServiceMonitor and PodMonitor scraping, Mimir remote write, cardinality filtering, alert rules, and GPU metrics.
- Logging Operations: Manage journal and Kubernetes event pipelines to Grafana Cloud Loki, including structured metadata, label strategy, RBAC isolation, and IPv6 readiness.
- Querying and Troubleshooting: Retrieve historical logs and metrics, validate credentials and namespaces, diagnose missing data, investigate Alloy resource usage, and resolve deployment issues across OSDC clusters.
Quick Start
Use the osdc-observability skill to diagnose why monitoring metrics or Kubernetes logs are missing from Grafana Cloud for a specified OSDC cluster.