venice-audio-speech

Convert text to speech with customizable TTS models and output formats.

130|15|Updated Apr 21, 2026
One-click install
npx skills add https://github.com/veniceai/skills --skill venice-audio-speech
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: venice-audio-speech
Source: https://github.com/veniceai/skills/tree/main/skills/venice-audio-speech
Command: npx skills add https://github.com/veniceai/skills --skill venice-audio-speech

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) components.

What problem does it solve?

This Skill transforms text into speech, enabling narrations, voice replies, or UI audio, and offers customizable options like TTS models, voices, and output formats.

Core Features & Use Cases

  • Text to Speech: Convert text to spoken words with a variety of TTS models and voices.
  • Customization: Adjust voice, format, streaming, emotion, and speed for tailored output.
  • Use Case: Create personalized audio messages, narrations for videos, or voice-activated systems.

Quick Start

Use the venice-audio-speech skill to generate speech from the text "Hello, welcome to Venice." in the 'tts-xai-v1' model and save it as 'hello.mp3'.

Frequently Asked Questions about venice-audio-speech

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert text to speech using customizable TTS models?

To convert text to speech, this Skill transforms your input text into spoken audio using customizable TTS models. You can generate narrations or voice replies and save the output in formats like MP3.

Can I adjust voice emotion and speed for text-to-speech generation?

Yes, text-to-speech generation supports extensive voice customization. You can adjust the voice, output format, streaming options, emotion, and speed to create tailored audio outputs for your specific use case.

Do I need authorization to access TTS models and select specific voices?

Yes, authorization is required to access TTS models and select specific voices. This ensures secure access to the available text-to-speech models and customizable voice options for generating audio.

What is the best way to generate streaming audio from text for UI applications?

The best way to generate streaming audio for UI applications is using this text-to-speech Skill. It transforms text into spoken words with streaming capabilities, making it ideal for voice replies and interactive UI audio.

Does the text-to-speech conversion support multiple output formats?

Yes, text-to-speech conversion supports a variety of output formats. You can customize the format of your generated speech, such as saving the audio output as an MP3 file for narrations or voice messages.