skill-creator

Build, refine, and benchmark automation skills through SKILL.md creation and evaluation loops.

Updated Feb 28, 2026
One-click install
npx skills add https://github.com/yuchenzhu-research/toefl-slang-master --skill skill-creator-yuchenzhu-research
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: skill-creator
Source: https://github.com/yuchenzhu-research/toefl-slang-master/tree/main/.agents/skills/skill-creator
Command: npx skills add https://github.com/yuchenzhu-research/toefl-slang-master --skill skill-creator-yuchenzhu-research

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Streamlines the end-to-end process of designing, testing, and iterating AI skills, helping teams quickly move from idea to measurable performance while maintaining quality and safety guards.

Core Features & Use Cases

  • Draft and refine SKILL.md to capture intent, workflow, and evaluation plans.
  • Run structured evals and benchmark loops to quantify triggering accuracy and outcomes.
  • Bundle and reuse scripts, references, and assets to scale skill development across projects.

Quick Start

Guides you from defining a new skill to drafting its SKILL.md and launching an initial evaluation workflow.

Frequently Asked Questions about skill-creator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I create and test AI skills for automation?

To create and test AI skills for automation, you can use a skill creator to capture intent, draft a SKILL.md file, run structured evaluations, and iterate on performance metrics to improve triggering accuracy across scenarios.

What is a SKILL.md file and how does it work?

A SKILL.md file captures the intent, workflow, and evaluation plans for an AI skill. It serves as the core definition document that guides how the skill operates, triggers, and measures performance across different automation scenarios.

How do I evaluate and benchmark AI skill performance?

You evaluate and benchmark AI skill performance by running structured evals and benchmark loops. This process quantifies triggering accuracy and outcomes, recording test prompts and metrics to enable repeatable improvement iterations.

Can I bundle scripts and references when building new skills?

Yes, you can bundle and reuse scripts, references, and assets when building new skills. This allows you to scale skill development across multiple projects while maintaining consistent quality and safety guards.

What's the best way to iterate on AI skill triggering accuracy?

The best way to iterate on AI skill triggering accuracy is using structured iteration loops. By recording test prompts, evaluating results, and benchmarking performance metrics, you can systematically refine the skill to improve outcomes.