nw-ab-critique-dimensions

Evaluate agent definitions against a formal quality framework with YAML review output.

Updated Mar 18, 2024
One-click install
npx skills add https://github.com/v1bh0r/precise-ledger-pro --skill nw-ab-critique-dimensions-v1bh0r
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: nw-ab-critique-dimensions
Source: https://github.com/v1bh0r/precise-ledger-pro/tree/main/nWave/skills/nw-ab-critique-dimensions
Command: npx skills add https://github.com/v1bh0r/precise-ledger-pro --skill nw-ab-critique-dimensions-v1bh0r

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

An evaluation framework that exposes and standardizes gaps in agent quality across template compliance, safety, and performance dimensions to streamline audits and remediation.

Core Features & Use Cases

  • Dimension-based evaluation covering Template Compliance, Size & Focus, Divergence Quality, Safety Implementation, Language & Tone, Examples Quality, Skill Loading, Token Efficiency, and Priority Validation
  • Structured review outputs with actionable remediation recommendations
  • Reusable framework for auditing Claude-like agents during development and governance

Quick Start

Review an agent using the nw-ab-critique-dimensions checklist and generate a structured assessment.

Frequently Asked Questions about nw-ab-critique-dimensions

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I evaluate agent quality across safety and performance dimensions?

To evaluate agent quality, apply a formal framework that assesses template compliance, safety implementation, and performance dimensions. This process identifies gaps and risks in Claude-like agents during development, generating structured YAML review outputs with actionable remediation recommendations.

What is a structured framework for auditing Claude-like agent definitions?

A structured auditing framework for Claude-like agents applies multi-dimension scoring across template compliance, divergence quality, and token efficiency. It extracts frontmatter-driven metadata to produce a formal assessment, yielding dimension ratings, identified issues, and a final verdict.

How do I systematically critique agent definitions for template compliance and safety?

Systematically critique agent definitions by applying a reusable evaluation checklist that covers template compliance, safety implementation, and language tone. This generates a structured assessment with dimension ratings and specific issues to streamline governance audits and remediation.

Does this agent quality review framework work for development and governance contexts?

Yes, the agent quality review framework works for both development and governance contexts. It evaluates Claude-like agents against a formal quality framework, exposing standardized gaps in template compliance, safety, and performance to streamline audits and guide remediation.

What dimensions should I review to identify risks in agent definitions?

To identify risks in agent definitions, review dimensions including size and focus, divergence quality, safety implementation, examples quality, skill loading, token efficiency, and priority validation. This multi-dimension scoring exposes hidden gaps and generates actionable remediation recommendations.

What is the output format when evaluating agent quality against a formal framework?

The output format when evaluating agent quality is a structured YAML review. It includes multi-dimension scoring, specific issues found during the evaluation, actionable remediation recommendations, and a final verdict based on the formal framework assessment.