principal-hierarchy-audit

Audit system prompts and plugin configurations against the Anthropic principal hierarchy.

1|Updated Oct 30, 2025
One-click install
npx skills add https://github.com/iamladi/cautious-computing-machine--primitives-plugin --skill principal-hierarchy-audit
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: principal-hierarchy-audit
Source: https://github.com/iamladi/cautious-computing-machine--primitives-plugin/tree/main/skills/principal-hierarchy-audit
Command: npx skills add https://github.com/iamladi/cautious-computing-machine--primitives-plugin --skill principal-hierarchy-audit

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Audits system prompts and plugin configurations against Anthropic's Constitutional principal hierarchy to identify instructions that conflict with Claude's training, attempt to weaponize Claude against users, violate inalienable user protections, or exceed operator permission boundaries.

Core Features & Use Cases

  • Detect identity deception, safety violations, and hard constraint breaches.
  • Classify issues by severity (RED/YELLOW/GREEN) with constitution references and provide actionable alternatives.
  • Generate structured audit reports with clear remediation guidance for operators.

Quick Start

Use a single sentence instruction to provide either a system prompt or a plugin file path to initiate the audit.

Frequently Asked Questions about principal-hierarchy-audit

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I audit system prompts against the Anthropic constitutional hierarchy?

Audit system prompts against the Anthropic constitutional hierarchy by providing a system prompt or plugin file path to generate a structured report. The report identifies compliance violations, classifies severities as RED, YELLOW, or GREEN, and returns recommended fixes.

What is a principal hierarchy audit for AI prompt compliance?

A principal hierarchy audit evaluates system prompts and plugin configurations to identify instructions that conflict with Claude's training, attempt to weaponize the AI, violate user protections, or exceed operator permission boundaries. It produces a structured machine-readable compliance report.

How do I check if my plugin configuration violates Anthropic safety policies?

Check plugin configurations for safety violations by running an audit that detects identity deception, hard constraint breaches, and operator boundary exceedances. The audit references specific constitutional principles and provides actionable compliant alternatives for each violation.

Do I need external API calls to run a prompt compliance audit?

No external API calls are required to run a prompt compliance audit. The tool operates locally by analyzing the provided system prompt or plugin file path and returning a self-contained machine-readable report with constitutional references and remediation guidance.

How do I classify prompt safety violations by severity?

Classify prompt safety violations by severity using an audit that generates RED, YELLOW, and GREEN classifications. The audit maps violations to implicated constitutional hierarchy levels and outputs a structured report with clear remediation guidance for operators.

Can I audit SKILL.md frontends for constitutional compliance?

Yes, you can audit SKILL.md frontends for constitutional compliance. The audit evaluates SKILL.md frontends, plugin prompts, and system prompts, producing a structured report with constitutional references and compliant alternatives for any identified violations.