voice-first-planning

Convert spoken transcripts into structured feature specifications and requirements.

37|8|Updated Mar 4, 2026
One-click install
npx skills add https://github.com/mlopscommunity/Coding-Agents-Conference-skills --skill voice-first-planning
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: voice-first-planning
Source: https://github.com/mlopscommunity/Coding-Agents-Conference-skills/tree/main/skills/voice-first-planning
Command: npx skills add https://github.com/mlopscommunity/Coding-Agents-Conference-skills --skill voice-first-planning

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill overcomes the limitations of typing for initial ideation and spec writing by allowing users to speak their thoughts naturally, capturing richer context and intent than traditional text input.

Core Features & Use Cases

  • Capture Raw Intent: Speak freely without self-editing to preserve unfiltered ideas.
  • Structure Unstructured Speech: LLMs process spoken transcripts into organized documents like feature specs or requirements.
  • Use Case: Start a new feature by dictating your initial thoughts, including the 'why', constraints, and edge cases, then have Claude Code structure it into a formal spec.

Quick Start

Use the voice-first-planning skill to structure the following spoken transcript into a feature spec with requirements, constraints, and open questions:

Frequently Asked Questions about voice-first-planning

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert speech to text for feature spec writing?

To convert speech to text for feature spec writing, you dictate raw thoughts and the Skill processes the transcript using LLMs to generate structured documents like feature specifications, requirements, and architectural decisions.

Why does typing limit initial ideation and spec writing?

Typing limits initial ideation and spec writing by enforcing self-editing that filters out context. Speaking thoughts naturally captures richer intent and unfiltered ideas for LLMs to structure into formal documents.

Can I dictate architectural decisions and have AI structure them into requirements?

Yes, you can dictate architectural decisions and have AI structure them into requirements. The Skill processes raw spoken transcripts to generate organized formal requirements, feature specs, and architectural documents.

What's the best way to capture unfiltered ideas for new feature ideation?

The best way to capture unfiltered ideas for new feature ideation is speaking freely without self-editing. The Skill preserves this raw context and uses LLMs to structure it into formal feature specifications.

Do I need a separate speech-to-text tool before structuring transcripts?

You need a speech-to-text tool to capture raw transcripts before this Skill processes them. The Skill leverages speech-to-text tools and LLMs to structure spoken language into formal feature specifications.