llm-evaluation
Evaluate Large Language Models with automated metrics and human evaluation frameworks.
npx skills add https://github.com/honysyang/skill-security-scanner --skill llm-evaluation-honysyang
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: llm-evaluation Source: https://github.com/honysyang/skill-security-scanner/tree/main/malicious-skills-research/llm-evaluation Command: npx skills add https://github.com/honysyang/skill-security-scanner --skill llm-evaluation-honysyang