tensorrt-llm
Optimize LLM inference on NVIDIA GPUs using TensorRT-LLM with quantization and batching.
npx skills add https://github.com/Orchestra-Research/AI-Research-SKILLs --skill tensorrt-llm-orchestra-research
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: tensorrt-llm Source: https://github.com/Orchestra-Research/AI-Research-SKILLs/tree/main/12-inference-serving/tensorrt-llm Command: npx skills add https://github.com/Orchestra-Research/AI-Research-SKILLs --skill tensorrt-llm-orchestra-research