What problem does it solve?
This Skill gives you a single, read-only snapshot of your local machine, SLURM GPU cluster, CPU server, and Matlantis jobs so you can quickly tell what is busy, queued, or idle without logging into each system separately.
Core Features & Use Cases
- Unified status view: Shows local CPU, RAM, and disk usage alongside remote cluster and server status in one tqdm-style display.
- SLURM monitoring: Summarizes your running and pending GPU jobs, GPU utilization by node and type, and queue pressure.
- CPU server monitoring: Highlights active compute processes and system load on a scheduler-less server.
- Matlantis job tracking: Lists papermill jobs, marks your own jobs, and reports progress only when the job has actually emitted a measurable value.
- Use Case: Ask for a quick HPC health check before starting new experiments, submitting jobs, or deciding whether to wait for resources.
Quick Start
Run the hpc-monitor skill to show my local machine, SLURM GPU cluster, CPU server, and Matlantis job status in one read-only snapshot.