What problem does it solve?
It diagnoses VMware vSphere incidents like CPU starvation, memory pressure cascades, storage latency cliffs, NUMA penalties, and snapshot accumulation with a structured, evidence-driven triage flow.
Core Features & Use Cases
- MCP-driven troubleshooting for vSphere signals: Queries Netdata via MCP to pull the key host/VM signals needed to identify the dominant failure archetype.
- Operator-playbook-based diagnostic tree: Applies the Netdata operator playbook logic to route an agent through the correct domain rules (CPU, Memory, Storage, Network, Availability/vCenter services, Hardware/Host Health, and others).
- Load-bearing verification contexts: Re-runs targeted MCP verification queries to confirm remediation actually moves signals back to expected ranges.
Quick Start
Use the troubleshoot-vmware-vsphere skill when investigating a VMware vSphere incident like widespread slowness or VM freezes by asking the agent to triage “a suspected storage latency cliff for host X over the last 15 minutes using Netdata via MCP.”