llama-cpp
Perform CPU-based LLM inference on non-NVIDIA hardware with GGUF quantization.
npx skills add https://github.com/overviewlabs/WHOX --skill llama-cpp-overviewlabs
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llama-cpp Source: https://github.com/overviewlabs/WHOX/tree/main/skills/mlops/inference/llama-cpp Command: npx skills add https://github.com/overviewlabs/WHOX --skill llama-cpp-overviewlabs