llama-cpp
Execute CPU-based LLM inference with GGUF quantization on Apple Silicon and non-NVIDIA GPUs.
npx skills add https://github.com/tangzheng202202/hermes-skills --skill llama-cpp-tangzheng202202
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llama-cpp Source: https://github.com/tangzheng202202/hermes-skills/tree/main/03-mlops/mlops/inference/llama-cpp Command: npx skills add https://github.com/tangzheng202202/hermes-skills --skill llama-cpp-tangzheng202202