agency-evaluation-criteria

Evaluate AI agency deliverables against quality criteria and generate evaluation-report.md.

Updated Mar 25, 2026
One-click install
npx skills add https://github.com/taewook486/anki_rag --skill agency-evaluation-criteria
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: agency-evaluation-criteria
Source: https://github.com/taewook486/anki_rag/tree/main/.claude/skills/agency-evaluation-criteria
Command: npx skills add https://github.com/taewook486/anki_rag --skill agency-evaluation-criteria

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Provides a standardized, skeptical evaluation framework and automated testing to ensure AI agency deliverables meet quality and compliance expectations.

Core Features & Use Cases

  • Weighted scoring across Design Quality, Originality, Completeness, and Functionality with explicit pass/fail criteria.
  • Automated testing via Playwright to verify deliverables against BRIEF, copy, and design-spec requirements.
  • Output generation of an evaluation-report.md containing overall scores, per-dimension evidence, defect lists with references, screenshots, and actionable improvements.
  • Applicable to agency projects including copywriting, UX/UI design, and implementation artifacts to enforce quality gates.

Quick Start

Provide your BRIEF, Original copy, and Original design-spec to generate a detailed evaluation report.

Frequently Asked Questions about agency-evaluation-criteria

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I evaluate AI agency project deliverables against quality criteria?

Evaluate AI agency deliverables by scoring them against weighted criteria like Design Quality, Originality, Completeness, and Functionality. This process uses input contracts such as BRIEF, Original copy, and design-spec files to generate a detailed evaluation report.

Can I use Playwright to automate quality testing for design and copy artifacts?

Yes, Playwright automates testing to verify deliverables against BRIEF, copy, and design-spec requirements. It captures screenshots and validates functionality to ensure AI agency outputs meet explicit pass/fail criteria.

What is included in an automated evaluation report for agency outputs?

An evaluation report includes an overall score, per-dimension scores with evidence, a defect list with file references and screenshots, and actionable improvement recommendations based on the input contracts.

How do I audit AI agency outputs for originality and completeness?

Audit AI agency outputs by applying a standardized evaluation framework that scores originality and completeness against the original copy and design specifications, producing a defect list with file references for any missing elements.

What inputs do I need to generate a quality evaluation report for an agency project?

You need to provide the BRIEF, Original copy.md, Original design-spec.md, and the built application artifacts. These inputs are used to evaluate the deliverables and generate the final evaluation-report.md.

Does the evaluation framework support different types of agency deliverables?

Yes, the framework is applicable to agency projects including copywriting, UX/UI design, and implementation artifacts, enforcing quality gates across various deliverable types using a weighted scoring system.