interpret-run
Parse run_outcome.json to diagnose AI benchmark failures and metrics.
npx skills add https://github.com/surus-lat/benchy --skill interpret-run
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: interpret-run Source: https://github.com/surus-lat/benchy/tree/main/.agent/skills/interpret-run Command: npx skills add https://github.com/surus-lat/benchy --skill interpret-run