llm-evaluation
Evaluate foundation models with BERTScore, ROUGE, and COMET metrics.
npx skills add https://github.com/hung-phan/ml-skills --skill llm-evaluation-hung-phan
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llm-evaluation Source: https://github.com/hung-phan/ml-skills/tree/main/skills/ml-review/references/ml-training/llm-evaluation Command: npx skills add https://github.com/hung-phan/ml-skills --skill llm-evaluation-hung-phan