tensorrt-llm
Optimize LLM inference with TensorRT-LLM on NVIDIA GPUs using quantization and multi-GPU scaling.
npx skills add https://github.com/BermudaLocals/hermes-agent-lite --skill tensorrt-llm-bermudalocals
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: tensorrt-llm Source: https://github.com/BermudaLocals/hermes-agent-lite/tree/main/optional-skills/mlops/tensorrt-llm Command: npx skills add https://github.com/BermudaLocals/hermes-agent-lite --skill tensorrt-llm-bermudalocals