scorer-reference

Catalogs 84 validated AIRT SDK scorers for AI red teaming vulnerability assessments.

11|1|Updated May 4, 2026
One-click install
npx skills add https://github.com/dreadnode/capabilities --skill scorer-reference
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: scorer-reference
Source: https://github.com/dreadnode/capabilities/tree/main/capabilities/ai-red-teaming/skills/scorer-reference
Command: npx skills add https://github.com/dreadnode/capabilities --skill scorer-reference

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill eliminates the time-consuming process of searching through scattered SDK documentation to find the correct scorer for AI red teaming tests, ensuring security teams can quickly select the right metrics for accurate vulnerability assessment.

Core Features & Use Cases

  • Comprehensive Scorer Catalog: 84 categorized scorers covering jailbreaks, agent security, MCP testing, exfiltration, and more, with exact valid names for use in the AIRT SDK scorers parameter.
  • Scenario Pairing Guide: Pre-built recommendations for which scorers to use for common attack types including system prompt extraction, credential leakage, tool abuse, and workflow manipulation.
  • Use Case: A red teamer testing a multi-agent customer service bot can instantly locate the relevant multi-agent security scorers to assess prompt infection and consensus poisoning risks without manual documentation searches.

Quick Start

Use the scorer-reference skill to look up the exact scorer names for testing MCP server tool description injection attacks.

Frequently Asked Questions about scorer-reference

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I find the right scorer for AI red teaming tests in the AIRT SDK?

To find the right AI red teaming scorer, use a complete catalog of 84 validated scorers categorized by rubric-based, pattern-based, agentic, and exfiltration metrics. This provides exact valid names and usage guidelines for the AIRT SDK `scorers` parameter to ensure accurate vulnerability detection.

What scorers are needed to test for system prompt extraction and credential leakage?

Testing system prompt extraction and credential leakage requires specific pattern-based and exfiltration scorers. A scenario pairing guide supplies pre-built recommendations matching these common attack types to the exact scorer names needed to assess data leakage risks during offensive security testing.

Can I assess multi-agent security risks like prompt infection using these scorers?

Yes, you can assess multi-agent security risks using dedicated multi-agent security scorers. These scorers evaluate prompt infection and consensus poisoning vulnerabilities in multi-agent workflows, allowing you to test multi-agent customer service bots without manual documentation searches.

Does this reference include scorers for MCP server tool description injection attacks?

Yes, this reference includes MCP security scorers specifically designed for testing MCP server tool description injection attacks. You can instantly locate the exact valid scorer names required to assess tool abuse and workflow manipulation risks within the AIRT SDK.

What categories of jailbreak scoring are available for vulnerability assessment?

Jailbreak scoring categories for vulnerability assessment include rubric-based and pattern-based scorers. These 84 categorized scorers cover jailbreaks, agent security, MCP testing, exfiltration, and reasoning, ensuring security teams can quickly select accurate metrics for testing jailbreak risks.