agent_evaluation_benchmarking
Evaluate AI agents across correctness, reliability, efficiency, and safety metrics.
npx skills add https://github.com/Renzo-Tognella/UniversalThingsForMyAgents --skill agent-evaluation-benchmarking
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: agent_evaluation_benchmarking Source: https://github.com/Renzo-Tognella/UniversalThingsForMyAgents/tree/main/skills/44_agent_evaluation_benchmarking Command: npx skills add https://github.com/Renzo-Tognella/UniversalThingsForMyAgents --skill agent-evaluation-benchmarking