skill-judge

Score SKILL.md files across eight dimensions using a 120-point rubric.

307|52|Updated Dec 30, 2025
One-click install
npx skills add https://github.com/shareAI-lab/shareAI-skills --skill skill-judge-shareai-lab
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: skill-judge
Source: https://github.com/shareAI-lab/shareAI-skills/tree/main/skills/skill-judge
Command: npx skills add https://github.com/shareAI-lab/shareAI-skills --skill skill-judge-shareai-lab

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill provides a structured approach to assess the quality, completeness, and compliance of Agent Skill design against official specifications and best practices. It helps reviewers identify gaps, reduce token waste, and produce actionable improvement recommendations.

Core Features & Use Cases

  • multidimensional evaluation across eight dimensions totaling 120 points, including Knowledge Delta, Mindset & Procedures, Anti-Pattern Quality, Specification Compliance, Progressive Disclosure, Freedom Calibration, Pattern Recognition, and Practical Usability.
  • an evaluation protocol that guides scoring, evidence gathering, and the generation of a formal Skill Evaluation Report.
  • use cases include auditing SKILL.md files, improving existing skill packages, and teaching new evaluators how to assess skills consistently.

Quick Start

Run the Skill Judge on the target skill by loading its SKILL.md and following the evaluation protocol.

Frequently Asked Questions about skill-judge

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I evaluate the quality of a SKILL.md file?

To evaluate a SKILL.md file, load the target skill and follow an evaluation protocol that scores frontmatter, body content, loading triggers, and references across eight dimensions to generate a structured report with actionable recommendations.

What is the best way to identify gaps in agent skill design?

The best way to identify gaps in agent skill design is by quantifying the knowledge delta between the skill and expert-level expectations, applying a 120-point scoring rubric across eight dimensions including specification compliance and practical usability.

How does skill evaluation scoring across multiple dimensions work?

Multidimensional skill evaluation scoring works by assessing eight specific dimensions totaling 120 points, including anti-pattern quality, progressive disclosure, and pattern recognition, to measure compliance and identify areas for improvement.

Can I use a structured protocol to audit existing agent skills?

Yes, you can use a structured evaluation protocol to audit existing agent skills by gathering evidence, scoring multidimensional criteria, and producing a formal Skill Evaluation Report to guide improvements and reduce token waste.

What are the limitations of evaluating agent skills with a rubric?

A limitation of evaluating agent skills with a rubric is that it primarily targets SKILL.md frontmatter, body content, loading triggers, and references, meaning it assesses design specifications rather than runtime execution performance.

When do I need to assess specification compliance for skill packages?

You need to assess specification compliance for skill packages when auditing existing files, improving skill quality, or teaching new evaluators how to measure design completeness against official best practices.