llama-cpp
Run local LLM inference on CPU and GPU with GGUF quantization.
npx skills add https://github.com/gqf2008/hermez-ai --skill llama-cpp-gqf2008
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llama-cpp Source: https://github.com/gqf2008/hermez-ai/tree/main/skills/mlops/inference/llama-cpp Command: npx skills add https://github.com/gqf2008/hermez-ai --skill llama-cpp-gqf2008