serving-llms-vllm
Deploy an OpenAI-compatible LLM inference server using vLLM with tensor parallelism and quantization.
npx skills add https://github.com/carterwayneskhizeine/hermes-agent-windows-R --skill serving-llms-vllm-carterwayneskhizeine
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: serving-llms-vllm Source: https://github.com/carterwayneskhizeine/hermes-agent-windows-R/tree/main/skills/mlops/inference/vllm Command: npx skills add https://github.com/carterwayneskhizeine/hermes-agent-windows-R --skill serving-llms-vllm-carterwayneskhizeine