evaluation-rubrics

Create and apply evaluation rubrics with criteria, scales, and descriptors.

142|20|Updated Oct 22, 2025
One-click install
npx skills add https://github.com/lyndonkl/claude --skill evaluation-rubrics
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: evaluation-rubrics
Source: https://github.com/lyndonkl/claude/tree/main/skills/evaluation-rubrics
Command: npx skills add https://github.com/lyndonkl/claude --skill evaluation-rubrics

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill helps you create and use evaluation rubrics to assess work objectively, ensuring consistency, fairness, and transparency in your evaluations.

Core Features & Use Cases

  • Rubric Design: Guides you through defining criteria, scales, and descriptors.
  • Consistent Evaluation: Enables standardized scoring across different reviewers and tasks.
  • Use Case: A team lead needs to evaluate code reviews from multiple engineers. They use this Skill to create a rubric with criteria like 'Code Clarity', 'Efficiency', and 'Test Coverage', ensuring all reviews are assessed against the same standards.

Quick Start

Use the evaluation-rubrics skill to create a rubric for evaluating technical blog posts.

Frequently Asked Questions about evaluation-rubrics

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I create an evaluation rubric for consistent grading and assessment?

An evaluation rubric defines specific criteria, scales, and performance descriptors to assess artifacts objectively. This Skill guides you through structuring those components to mitigate subjective bias and ensure consistent, transparent grading across different reviewers.

What is the best way to standardize scoring criteria for code reviews?

Standardizing scoring criteria for code reviews requires defining explicit performance descriptors for dimensions like clarity and test coverage. This Skill helps establish those objective quality standards so all engineering reviews are evaluated against the same scale.

How does defining performance descriptors improve inter-rater reliability?

Defining performance descriptors improves inter-rater reliability by giving multiple reviewers a standardized scale to anchor their judgments. This Skill facilitates creating those detailed rubrics, reducing subjective bias during design critiques and writing evaluations.

Can I use this rubric design process for academic grading and design critiques?

Yes, the rubric design process applies to academic grading, design critiques, and writing evaluations. This Skill supports defining assessment criteria and scales for various artifacts, ensuring objective quality standards across different evaluation contexts.

What should I do if my evaluation rubric is not mitigating subjective bias effectively?

If your evaluation rubric is not mitigating subjective bias, your performance descriptors may lack specific scales. This Skill helps refine criteria definitions and descriptor clarity to improve inter-rater reliability and objective assessment consistency.