simpo-training
Optimize LLM preference alignment without a reference model using PyTorch and YAML.
npx skills add https://github.com/adm-humanerd/drewgent --skill simpo-training-adm-humanerd
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: simpo-training Source: https://github.com/adm-humanerd/drewgent/tree/main/optional-skills/mlops/simpo Command: npx skills add https://github.com/adm-humanerd/drewgent --skill simpo-training-adm-humanerd