llama-cpp
Run LLM inference on CPU, Apple Silicon, and non-NVIDIA GPUs with GGUF quantization.
npx skills add https://github.com/attentiondotnet/hermes-agent --skill llama-cpp-attentiondotnet
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llama-cpp Source: https://github.com/attentiondotnet/hermes-agent/tree/main/skills/mlops/inference/llama-cpp Command: npx skills add https://github.com/attentiondotnet/hermes-agent --skill llama-cpp-attentiondotnet