tensorrt-llm
Optimize LLM inference on NVIDIA GPUs with TensorRT-LLM quantization and multi-GPU scaling.
npx skills add https://github.com/gqf2008/hermez-ai --skill tensorrt-llm-gqf2008
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: tensorrt-llm Source: https://github.com/gqf2008/hermez-ai/tree/main/skills/mlops/tensorrt-llm Command: npx skills add https://github.com/gqf2008/hermez-ai --skill tensorrt-llm-gqf2008