optimizing-attention-flash
Implement Flash Attention to speed up Transformer inference and reduce memory use.
npx skills add https://github.com/blueskies1818/hermesALIone --skill optimizing-attention-flash-blueskies1818
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: optimizing-attention-flash Source: https://github.com/blueskies1818/hermesALIone/tree/main/Agent/optional-skills/mlops/flash-attention Command: npx skills add https://github.com/blueskies1818/hermesALIone --skill optimizing-attention-flash-blueskies1818