model-eval-benchmark
Evaluate brain-app JSONL agent tasks across models to produce comparable reports.
npx skills add https://github.com/cirne/brain-app --skill model-eval-benchmark
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: model-eval-benchmark Source: https://github.com/cirne/brain-app/tree/main/.cursor/skills/model-eval-benchmark Command: npx skills add https://github.com/cirne/brain-app --skill model-eval-benchmark