sadd:judge

Runs a two-phase evaluation producing an evidence-backed report with YAML specification.

2|Updated Mar 30, 2026
One-click install
npx skills add https://github.com/fockus/claude-skill-build --skill sadd-judge-fockus
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: sadd:judge
Source: https://github.com/fockus/claude-skill-build/tree/main/skills/sadd-judge
Command: npx skills add https://github.com/fockus/claude-skill-build --skill sadd-judge-fockus

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Two-phase evaluation pipeline to assess and report on work produced in a conversation: a meta-judge generates tailored evaluation criteria, followed by a judge applying those criteria with isolated context, producing a structured, evidence-backed report without modifying the original work.

Core Features & Use Cases

  • Context extraction and scope definition to ensure focused evaluation.
  • Meta-judge generates rubrics and scoring criteria tailored to the artifact.
  • Judge applies criteria with fresh context and requires citations for every score.
  • Phase 4 processing validates the evaluation and presents the report without modifying the work.

Quick Start

Provide the work context to sadd:judge and initiate the two-phase evaluation workflow.

Frequently Asked Questions about sadd:judge

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I evaluate work artifacts with tailored rubrics and evidence-backed reporting?

To evaluate work artifacts with tailored rubrics, you initiate a two-phase evaluation pipeline that generates custom scoring criteria and applies them with isolated context to produce a structured, evidence-backed report without modifying the original work.

What is a two-phase meta-judge evaluation pipeline?

A two-phase meta-judge evaluation pipeline is a process where a meta-judge first extracts context to produce a YAML evaluation specification, followed by a judge applying those tailored criteria with fresh context to generate an evidence-backed report.

How do I generate custom evaluation criteria for conversation artifacts?

You generate custom evaluation criteria by providing work context to the system, which dispatches a meta-judge to produce a tailored YAML evaluation specification that defines the rubrics and scoring rules for the artifact.

Does the evaluation process modify the original work being assessed?

The evaluation process does not modify the original work, because the system enforces strict context isolation during the evaluation phases and validates the results before presenting a structured report.

Can I use isolated context to prevent bias during artifact evaluation?

Yes, you can use isolated context to prevent bias, because the system applies freshly generated rubrics to the work artifact with strict context isolation and enforces evidence requirements for every score to mitigate bias.

What are the limitations of automated rubric-based artifact evaluation?

A limitation of automated rubric-based artifact evaluation is that it requires sufficient work context for the meta-judge to generate meaningful criteria, and the judge must enforce strict evidence citations for every score to ensure validity.