skill-benchmark

Generate standardized benchmark reports for reusable agent skill packages.

Updated Apr 27, 2026
One-click install
npx skills add https://github.com/ginmp8/rhapsodia --skill skill-benchmark-ginmp8
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: skill-benchmark
Source: https://github.com/ginmp8/rhapsodia/tree/main/skills/skill-benchmark
Command: npx skills add https://github.com/ginmp8/rhapsodia --skill skill-benchmark-ginmp8

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) and assets (resource) components.

What problem does it solve?

Benchmarks provide evidence-based evaluation of reusable agent skill packages by generating standardized benchmark reports.

Core Features & Use Cases

  • Generate standardized benchmark reports for reusable skill packages and benchmark results.
  • Produce a structured scorecard, gates, inventory, scenario status, risks, improvements, and verdict for publish readiness.
  • Compare versions, validate reports, and assess readiness based on supplied scenario results.

Quick Start

Place the target skill folder under skills/ and run the benchmark tool to generate a canonical report.

Frequently Asked Questions about skill-benchmark

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I benchmark reusable agent skill packages for publish readiness?

Benchmarks provide evidence-based evaluation of reusable agent skill packages by generating standardized benchmark reports that validate structure and surface actionable findings without mutating the target package.

What is included in a standardized benchmark report for skill packages?

A standardized benchmark report includes a structured scorecard, gates, inventory, scenario status, risks, improvements, and a final verdict for publish readiness based on supplied scenario results.

How do I validate and compare versions of skill packages?

You can validate reports and compare versions of skill packages by applying the benchmark tool to assess readiness based on supplied scenario results, producing a scorecard and verdict for each version.

Does the benchmark tool modify the target skill package during evaluation?

No, the benchmark tool avoids mutations to the target package, strictly validating structure and coordinating with scenario results to produce an evidence-based evaluation.

Can I assess target skills without supplying scenario results?

The benchmark tool coordinates with supplied scenario results to assess readiness, producing scenario status and gates, so supplying scenario results is necessary for a complete publish readiness verdict.