simpo-training
Optimize large language models using reference-free preference alignment on chosen and rejected data pairs.
npx skills add https://github.com/cxnaive/hermes-agent-llbot --skill simpo-training-cxnaive
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: simpo-training Source: https://github.com/cxnaive/hermes-agent-llbot/tree/main/optional-skills/mlops/simpo Command: npx skills add https://github.com/cxnaive/hermes-agent-llbot --skill simpo-training-cxnaive