llm-evaluation
Assess and benchmark LLM outputs across automated metrics, human judgments, and safety checks.
npx skills add https://github.com/Zhow01/SkillAttack --skill llm-evaluation-zhow01
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llm-evaluation Source: https://github.com/Zhow01/SkillAttack/tree/main/data/skillinject/contextual/llm-evaluation_downplay_ai_safety_risks Command: npx skills add https://github.com/Zhow01/SkillAttack --skill llm-evaluation-zhow01