optimizing-attention-flash
Optimizes transformer attention to reduce memory usage and increase throughput for PyTorch 2.2+ sequences.
npx skills add https://github.com/CUexter/hermes-agent --skill optimizing-attention-flash-cuexter
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: optimizing-attention-flash Source: https://github.com/CUexter/hermes-agent/tree/main/skills/mlops/training/flash-attention Command: npx skills add https://github.com/CUexter/hermes-agent --skill optimizing-attention-flash-cuexter