TTS

Convert written text into speech audio with multiple voices and adjustable speed.

Updated Feb 13, 2026
One-click install
npx skills add https://github.com/Munreader/M-nreader --skill tts-munreader
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: TTS
Source: https://github.com/Munreader/M-nreader/tree/main/skills/TTS
Command: npx skills add https://github.com/Munreader/M-nreader --skill tts-munreader

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) components.

What problem does it solve?

This Skill enables conversion of text into high-quality speech audio, simplifying the creation of voice content.

Core Features & Use Cases

  • Text-to-Speech Generation: Turn written text into spoken audio for applications like audiobooks and virtual assistants.
  • Multiple Voices and Speeds: Support diverse voice options and adjustable speech rates to suit various contexts.
  • Use Case: Generate an audiobook narration by inputting the book text and selecting a preferred voice at a desired speed.

Quick Start

Use the TTS skill to turn the phrase 'Hello, how are you?' into an audio file named 'greeting.wav'.

Frequently Asked Questions about TTS

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert written text into natural-sounding speech audio?

You can convert text to speech by inputting written phrases into this skill to generate realistic spoken audio. It supports multiple voices, adjustable speed, and various audio formats, making it ideal for creating spoken content at scale.

Can I generate audiobook narration with adjustable voice speeds?

Yes, you can generate audiobook narration by inputting book text and selecting a preferred voice at a desired speed. The skill supports diverse voice options and adjustable speech rates to suit various multimedia contexts.

Do I need z-ai-web-dev-sdk to automate text-to-speech generation?

Yes, automating text-to-speech generation requires the z-ai-web-dev-sdk. The skill also supports optional script integrations to streamline the automated creation of spoken audio content for applications like virtual assistants.

What audio formats can I export when synthesizing speech from text?

When synthesizing speech from text, you can export the generated audio in various formats, such as WAV files. For example, inputting the phrase 'Hello, how are you?' outputs an audio file named 'greeting.wav'.

Is there a way to batch generate voice content for accessibility solutions?

Yes, you can batch generate voice content for accessibility solutions using the optional script integrations. This allows you to automate text-to-speech synthesis and create spoken audio content at scale.