advanced-evaluation
Implement LLM-as-a-Judge direct scoring and pairwise comparison for model output evaluation.
npx skills add https://github.com/bthillerup/bens-garage-session-2 --skill advanced-evaluation-bthillerup
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: advanced-evaluation Source: https://github.com/bthillerup/bens-garage-session-2/tree/main/.github/skills/advanced-evaluation Command: npx skills add https://github.com/bthillerup/bens-garage-session-2 --skill advanced-evaluation-bthillerup