What problem does it solve?
It helps you quickly triage and remediate ZFS performance and health incidents by identifying the most likely failure archetype and verifying it with Netdata-observed signals.
Core Features & Use Cases
- Structured ZFS triage for core failure archetypes: pool health degradation, capacity-fragmentation cliff, TXG sync hang, ARC memory starvation, silent data corruption, and ZIL/SLOG bottlenecks.
- Netdata + MCP query-driven diagnostics: pulls the relevant MCP signals for pool health, per-vdev state and errors, I/O operations/latency/queue depth, and internal-state signals, then routes you through the operator playbook-style decision tree.
- Signal verification before and after remediation: re-runs the same MCP queries to confirm the system returns to expected ranges rather than stopping at the first anomaly.
Quick Start
Use this skill when you see a ZFS pool or service behaving abnormally, and ask an AI coding agent to troubleshoot by querying Netdata via MCP and following the remediation path for the matching failure archetype.