skill-creator

Draft, test, and benchmark AI skill instructions with automated evaluation loops.

4|Updated Jun 10, 2026
One-click install
npx skills add https://github.com/mymx2/foreman --skill skill-creator-mymx2
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: skill-creator
Source: https://github.com/mymx2/foreman/tree/main/.qoder/skills.collected/skills/skill-creator
Command: npx skills add https://github.com/mymx2/foreman --skill skill-creator-mymx2

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) and assets (resource) components.

What problem does it solve?

This skill streamlines the complex process of developing, evaluating, and refining AI-agent skills, ensuring they are robust, accurate, and performant before deployment.

Core Features & Use Cases

  • Iterative Development: Provides a structured loop for drafting, testing, and refining skill instructions based on real-world performance.
  • Quantitative Benchmarking: Automates the creation of test cases, execution of baseline comparisons, and generation of performance metrics.
  • Trigger Optimization: Includes a specialized loop to refine skill descriptions, ensuring the AI triggers the skill only when appropriate.
  • Use Case: If you are building a custom skill for data analysis, use this tool to run it against a set of test prompts, compare the results against a baseline, and automatically improve the skill's instructions based on the feedback.

Quick Start

Use the skill-creator to draft a new skill for summarizing technical documentation and set up the initial test cases.

Frequently Asked Questions about skill-creator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate testing and benchmarking for AI agent skills?

Automating AI agent skill testing involves using iterative evaluation loops to execute baseline comparisons and generate quantitative performance metrics. This skill automates test case creation and behavior analysis to refine instructions.

What is the best way to optimize AI skill descriptions for accurate triggering?

Optimizing AI skill descriptions for triggering requires a specialized refinement loop that evaluates relevance across diverse user prompts. This skill adjusts descriptions quantitatively to ensure accurate activation only when appropriate.

How do I set up an iterative development loop for custom AI skills?

Setting up an iterative development loop for custom AI skills requires drafting, testing, and refining instructions based on real-world performance. This skill facilitates this cycle through automated evaluations and feedback integration.

Can I run quantitative performance comparisons for my AI development prompts?

Yes, you can run quantitative performance comparisons for AI development prompts. This skill automates baseline comparisons and generates metrics by executing test cases against your drafted instructions.

Why does my custom AI skill trigger incorrectly on irrelevant user prompts?

Custom AI skills trigger incorrectly when descriptions lack optimization for diverse user prompts. This skill provides a specialized evaluation loop to refine descriptions and improve triggering accuracy.

Does skill-creator support automated evaluation of technical documentation summarization?

Yes, skill-creator supports automated evaluation of technical documentation summarization. You can draft a summarization skill, set up test cases, and run automated evaluations to compare results against a baseline.