prompt-benchmark
Evaluate AI prompt effectiveness on standardized benchmarks like MATH and GSM8K.
npx skills add https://github.com/HermeticOrmus/hermetic-claude --skill prompt-benchmark-hermeticormus
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: prompt-benchmark Source: https://github.com/HermeticOrmus/hermetic-claude/tree/main/claude/skills/prompt-benchmark Command: npx skills add https://github.com/HermeticOrmus/hermetic-claude --skill prompt-benchmark-hermeticormus