text-to-speech

Convert text into speech audio using HeyGen's Starfish TTS model.

Updated May 23, 2026
One-click install
npx skills add https://github.com/xingBaGan/FANovelist --skill text-to-speech-xingbagan
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: text-to-speech
Source: https://github.com/xingBaGan/FANovelist/tree/main/src/openharness/openmontage/.claude/skills/text-to-speech
Command: npx skills add https://github.com/xingBaGan/FANovelist --skill text-to-speech-xingbagan

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires mcp__heygen__*, and includes scripts (resource) components.

What problem does it solve?

This Skill solves the need for generating speech audio from text without manual voiceover work, providing flexibility in voice selection, speed, and pitch.

Core Features & Use Cases

  • Text to Speech Conversion: Convert any given text into high-quality speech audio using a wide range of voice options.
  • Voice Customization: Select from various languages and genders to match the desired voice characteristics.
  • Speed and Pitch Control: Adjust the speed and pitch of the generated speech for optimal listening experience.
  • Applications: Ideal for creating voiceovers, narrations, podcasts, and more, as well as for integrating with HeyGen's audio endpoints.

Quick Start

Use the text-to-speech skill to generate speech audio from the text "Hello! Welcome to our product demo."

Frequently Asked Questions about text-to-speech

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate voiceovers from text without manual recording?

It converts text into speech audio using the HeyGen Starfish TTS model, supporting multiple languages and genders so you can match the voice characteristics to your desired audience.

Can I adjust the speed and pitch of generated speech audio?

Yes, you can adjust the speed and pitch of generated speech audio to achieve your desired listening experience and match the voiceover to your content style.

Do I need an API key to use HeyGen for text-to-speech conversion?

Yes, you need a HEYGEN_API_KEY for authentication to use the HeyGen Starfish TTS model and access its voice generation capabilities through MCP tools.

What is the best way to create multilingual narrations from written text?

Using a TTS model that supports various languages and genders is the best way to create personalized narrations, enabling you to select voice characteristics that match your content requirements.

Does this text-to-speech skill work for creating podcast audio?

Yes, text-to-speech conversion is ideal for creating podcast audio and narrations, providing flexibility in voice selection without requiring manual voiceover work.