serving-llms-vllm
Configure vLLM servers with OpenAI-compatible endpoints, quantization options, and tensor parallelism for scalable LLM deployments on GPU infrastructure.
npx skills add https://github.com/tangzheng202202/hermes-skills --skill serving-llms-vllm-tangzheng202202
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: serving-llms-vllm Source: https://github.com/tangzheng202202/hermes-skills/tree/main/03-mlops/mlops/inference/vllm Command: npx skills add https://github.com/tangzheng202202/hermes-skills --skill serving-llms-vllm-tangzheng202202