skill-creator

Create, test, evaluate, and refine Claude skills through a structured workflow.

5|3|Updated Jan 22, 2026
One-click install
npx skills add https://github.com/hackutd/harp --skill skill-creator-hackutd
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: skill-creator
Source: https://github.com/hackutd/harp/tree/main/.claude/skills/skill-creator
Command: npx skills add https://github.com/hackutd/harp --skill skill-creator-hackutd

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill provides a structured process for designing, testing, evaluating, and improving Claude skills instead of relying on ad hoc instructions or subjective iteration.

Core Features & Use Cases

  • Skill Creation: Capture user intent, define triggering conditions, specify output formats, and draft complete SKILL.md files.
  • Performance Evaluation: Create test cases, run skill-enabled and baseline comparisons, grade assertions, and analyze benchmark results.
  • Iterative Improvement: Review qualitative feedback, identify recurring weaknesses, refine workflows, and compare successive skill versions.
  • Description Optimization: Generate trigger evaluations and optimize skill descriptions for more accurate activation.
  • Use Case: Use this Skill to turn a repeatable document-processing, coding, or research workflow into a tested and discoverable Claude skill.

Quick Start

Use the skill-creator skill to design or improve a Claude skill, create realistic evaluations, benchmark its performance, and optimize its triggering description.

Frequently Asked Questions about skill-creator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I create and test Claude skills with a structured workflow?

To create and test Claude skills, use a structured workflow to capture intent, define trigger conditions, draft SKILL.md files, run baseline comparisons, and grade assertions for iterative improvement. This replaces ad hoc instructions with evaluated skill designs.

What is the best way to evaluate and benchmark prompt engineering skills?

Evaluating prompt engineering skills requires creating test cases, running skill-enabled and baseline comparisons, grading assertions, and aggregating benchmark results to identify performance gaps and recurring weaknesses for iterative refinement.

How do I optimize skill descriptions for more accurate activation?

Optimizing skill descriptions involves generating trigger evaluations to test activation accuracy, refining the descriptive metadata, and comparing successive versions to ensure the skill activates reliably under the intended conditions.

Can I turn a repeatable document-processing or research workflow into a Claude skill?

Yes, you can convert repeatable document-processing, coding, or research workflows into tested and discoverable Claude skills by defining triggering conditions, specifying output formats, and drafting complete SKILL.md files.

Why does my Claude skill fail to activate or trigger consistently?

Inconsistent skill activation often stems from poorly optimized trigger descriptions. Running trigger evaluations and iterating on the activation description helps ensure the skill fires reliably for the intended workflows.