vllm-omni-quantization
Quantizes vLLM-Omni models using AWQ, GPTQ, or FP8 for reduced memory and faster throughput.
npx skills add https://github.com/hsliuustc0106/vllm-omni-skills --skill vllm-omni-quantization
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: vllm-omni-quantization Source: https://github.com/hsliuustc0106/vllm-omni-skills/tree/main/skills/vllm-omni-quantization Command: npx skills add https://github.com/hsliuustc0106/vllm-omni-skills --skill vllm-omni-quantization