k8s-platform-operations

Coordinate Kubernetes platform operations for health checks, incident response, and capacity planning.

18|2|Updated Jan 30, 2026
One-click install
npx skills add https://github.com/foxj77/claude-code-skills --skill k8s-platform-operations
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: k8s-platform-operations
Source: https://github.com/foxj77/claude-code-skills/tree/main/k8s-platform-operations
Command: npx skills add https://github.com/foxj77/claude-code-skills --skill k8s-platform-operations

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

Keeps Kubernetes platforms healthy and responsive by providing structured guidance for monitoring, incident response, capacity planning, maintenance, backups, and runbooks.

Core Features & Use Cases

  • Operational playbooks for cluster health checks, incident response, capacity planning, and routine runbooks.
  • Comprehensive maintenance workflows, backup and recovery procedures, and on-call handoff guidance.
  • Clear escalation paths and runbooks that standardize actions across multi-tenant Kubernetes environments.

Quick Start

Review the Health Monitoring and Incident Response sections to perform an end-to-end platform health check and basic remediation steps.

Frequently Asked Questions about k8s-platform-operations

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I manage Kubernetes platform operations and incident response?

Kubernetes platform operations are managed using structured runbooks that coordinate cluster health checks, incident remediation, and maintenance workflows. These procedures standardize on-call handoffs and escalation paths across environments.

What is the best way to perform Kubernetes health checks and capacity planning?

Kubernetes health checks and capacity planning are executed through operational playbooks. These provide structured procedures to monitor cluster health and proactively plan resource capacity for multi-tenant environments.

Can I use kubectl and Velero for cluster backups and recovery?

Yes, kubectl and Velero integrate with the command-focused guidelines for backups and recovery. The runbooks provide structured workflows to execute backup operations and restore procedures across Kubernetes environments.

Does this approach work for multi-tenant Kubernetes environments?

Yes, the runbooks and escalation workflows are designed to standardize actions across multi-tenant Kubernetes environments. They provide clear maintenance workflows and on-call handoff guidance tailored for multi-tenant operations.

How do I standardize on-call handoffs and escalation paths for cluster maintenance?

On-call handoffs and escalation paths are standardized using comprehensive maintenance workflows and operational playbooks. These define clear escalation procedures and runbook execution steps for routine Kubernetes maintenance.