minimax-tts

Convert input text into speech audio using MiniMax's T2A API.

2|1|Updated Feb 16, 2026
One-click install
npx skills add https://github.com/desirecore/market --skill minimax-tts-desirecore
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: minimax-tts
Source: https://github.com/desirecore/market/tree/main/skills/minimax-tts
Command: npx skills add https://github.com/desirecore/market --skill minimax-tts-desirecore

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill enables converting text into natural speech audio, facilitating voice synthesis for various multimedia applications.

Core Features & Use Cases

  • Multiple Voice Styles: Supports a variety of voice personas suited for narration, narration, or character voices.
  • Emotional Control & Cloning: Allows adjustment of tone and voice cloning for personalized sound outputs.
  • Use Case: Generate a narrated audiobook from a script by converting text to audio clips that can be assembled into a complete recording.

Quick Start

Provide the text you want to turn into speech to instantly generate an audio file.

Frequently Asked Questions about minimax-tts

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert text to speech audio for multimedia voiceovers?

To convert text to speech for multimedia voiceovers, you can use this Skill to process input text through the MiniMax T2A API, generating high-quality speech audio suitable for immediate use in multimedia projects.

Can I adjust the emotional tone and voice style for text-to-speech synthesis?

Yes, text-to-speech synthesis supports adjusting emotional tone and selecting from multiple voice personas, allowing you to generate narration or character voices with personalized emotional control.

What do I need to generate a narrated audiobook from a script?

To generate a narrated audiobook from a script, you need to provide the text input which is instantly converted into audio clips that can be assembled into a complete voice recording.

Does voice cloning work for personalized text-to-speech outputs?

Voice cloning works for personalized text-to-speech outputs by allowing tone adjustment and voice synthesis customization, resulting in tailored audio files suited for specific sound profiles.

What's the best way to automate narration for accessibility needs?

The best way to automate narration for accessibility needs is processing text through the T2A API, which provides immediate speech synthesis to generate accessible audio files from written content.