llama-cpp
Enable CPU-based LLM inference with GGUF quantization and OpenAI-compatible server endpoints.
npx skills add https://github.com/xiaoquqi/hermes-agent-skills --skill llama-cpp-xiaoquqi
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llama-cpp Source: https://github.com/xiaoquqi/hermes-agent-skills/tree/main/mlops/inference/llama-cpp Command: npx skills add https://github.com/xiaoquqi/hermes-agent-skills --skill llama-cpp-xiaoquqi