skill-creator

Automate drafting, testing, and refining Claude skills with evaluation loops.

5|Updated Feb 6, 2022
One-click install
npx skills add https://github.com/christofferbergj/dotfiles --skill skill-creator-christofferbergj
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: skill-creator
Source: https://github.com/christofferbergj/dotfiles/tree/main/.agents/skills/skill-creator
Command: npx skills add https://github.com/christofferbergj/dotfiles --skill skill-creator-christofferbergj

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires anthropic, yaml, and includes scripts (resource) and references (resource) components.

What problem does it solve?

The Skill Creator automates the end-to-end process of drafting Claude skills, running evaluation loops, and iterating on descriptions, reducing manual trial-and-error and speeding up skill quality improvements.

Core Features & Use Cases

  • Draft and structure SKILL.md: Generate a solid, discovery-friendly frontmatter and body that guide activation and usage.
  • Automated evaluation loop: Run trigger evaluations, collect results, and iterate on prompts, prompts history, and descriptions.
  • Benchmark and review integration: Produce evaluation reports, benchmarks, and reviewer-ready notes to inform improvements and stakeholder decisions.

Quick Start

Run the evaluation loop with your eval-set and skill-path to begin drafting, testing, and refining Claude skills.

Frequently Asked Questions about skill-creator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate Claude skill creation and evaluation loops?

Automating Claude skill creation involves generating a SKILL.md file, running trigger evaluations, and iterating on prompts. This tool applies Claude-based tooling to run end-to-end cycles from draft to improved description.

What is an evaluation loop for prompt engineering and skill refinement?

An evaluation loop for prompt engineering tests skill descriptions against eval-sets, collects results, and iteratively refines prompts. It reduces manual trial-and-error by automating trigger evaluations and producing benchmark reports.

How do I generate discovery-friendly frontmatter and structure for SKILL.md files?

Generating discovery-friendly frontmatter and structure for SKILL.md files requires drafting a solid body that guides activation and usage. This approach automates structuring to speed up skill quality improvements.

Can I use anthropic and yaml dependencies to benchmark Claude skills?

Yes, you can use anthropic and yaml dependencies to benchmark Claude skills. The tooling produces evaluation reports, benchmarks, and reviewer-ready notes to inform improvements and stakeholder decisions.

What's the best way to measure Claude skill performance across multiple evals?

The best way to measure Claude skill performance across multiple evals is running an automated evaluation loop with your eval-set and skill-path. This collects results and produces benchmark reports to optimize descriptions.