grpo-rl-training
Implement GRPO-based RL fine-tuning for language models with TRL.
npx skills add https://github.com/jacardl/New-Radar --skill grpo-rl-training-jacardl
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: grpo-rl-training Source: https://github.com/jacardl/New-Radar/tree/main/backend/frameworks/hermes-agent/skills/mlops/training/grpo-rl-training Command: npx skills add https://github.com/jacardl/New-Radar --skill grpo-rl-training-jacardl