inference-optimization
Optimize AI model inference speed with quantization, speculative decoding, KV caching, and batching.
npx skills add https://github.com/doanchienthangdev/omgkit --skill inference-optimization
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: inference-optimization Source: https://github.com/doanchienthangdev/omgkit/tree/main/plugin/skills/ai-engineering/inference-optimization Command: npx skills add https://github.com/doanchienthangdev/omgkit --skill inference-optimization