tensorrt-llm
Optimize LLM inference throughput and latency on NVIDIA GPUs with TensorRT-LLM.
npx skills add https://github.com/arsity/scholar-tools --skill tensorrt-llm-arsity
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: tensorrt-llm Source: https://github.com/arsity/scholar-tools/tree/main/vendor/ai-research-skills/12-inference-serving/tensorrt-llm Command: npx skills add https://github.com/arsity/scholar-tools --skill tensorrt-llm-arsity