skill-creator

Automate the end-to-end workflow of creating, evaluating, and improving Claude skills.

Updated Jan 28, 2026
One-click install
npx skills add https://github.com/xjiang77/agent-wisp --skill skill-creator-xjiang77
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: skill-creator
Source: https://github.com/xjiang77/agent-wisp/tree/main/tools/setup-toolkit/skills/skill-creator
Command: npx skills add https://github.com/xjiang77/agent-wisp --skill skill-creator-xjiang77

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

The Skill Creator helps teams design, test, and iteratively improve Claude skills, reducing manual trial-and-error in building capable AI helpers.

Core Features & Use Cases

  • End-to-end skill development: draft, test prompts, run evals, and refine based on feedback.
  • Iterative evaluation: orchestrates evaluation, grading, comparison, and analysis to identify best versions.
  • Flexible workflows: supports creation, improvement, and benchmarking modes with task tracking and version history.

Quick Start

Initialize the skill workspace and draft the SKILL.md, then run the iterative development loop to draft, test, and refine the skill.

Frequently Asked Questions about skill-creator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I create and evaluate Claude skills automatically?

To create and evaluate Claude skills, you can automate the end-to-end workflow of drafting prompts, running evals, analyzing test results, and iterating until performance goals are met across multiple versions.

What is the best way to iteratively improve AI prompt performance?

The best way to iteratively improve AI prompt performance is by running a structured evaluation loop that grades test results, compares versioned iterations, and applies actionable feedback loops to refine the drafts.

How does iterative skill evaluation work for Claude plugins?

Iterative skill evaluation works by orchestrating evaluation milestones, grading test outcomes, comparing versioned iterations, and analyzing results to identify the best performing skill versions for Claude plugins.

Can I benchmark different versions of a skill during development?

Yes, you can benchmark different versions of a skill during development by using the benchmarking mode, which supports task tracking and version history to compare and identify the best iterations against your goals.

Do I need any external plugins to run skill evaluation workflows?

No external plugins are required to run skill evaluation workflows, as the coordinator operates independently without dependencies to guide the creation, improvement, and benchmarking of your skills.

Why does my skill development workflow require manual trial-and-error?

Your skill development requires manual trial-and-error if you lack an automated iterative evaluation loop, which orchestrates drafting, testing, grading, and refinement to systematically reduce manual effort and improve quality.