Voice Synthesis

Generate production voice audio from SSML scripts via Google Cloud Text-to-Speech.

Updated Aug 27, 2026
One-click install
npx skills add https://github.com/randysalars/dreamweaving --skill voice-synthesis
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: Voice Synthesis
Source: https://github.com/randysalars/dreamweaving/tree/main/.claude/skills/tier3-production/voice-synthesis
Command: npx skills add https://github.com/randysalars/dreamweaving --skill voice-synthesis

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-cloud-texttospeech.

What problem does it solve?

Converts SSML scripts into high-quality hypnotic voice audio using Google Cloud Text-to-Speech with psychoacoustic enhancements, reducing manual production time.

Core Features & Use Cases

  • Production-grade voice generation: Generates voice audio from SSML with production defaults, enhancements, and chunking for long scripts.
  • Voice standardization: Applies a defined production voice (en-US-Neural2-H) and preset parameters for consistent output.
  • Use Case: Create immersive hypnotic sessions by turning SSML into ready-to-mix audio.

Quick Start

Execute a single command to generate the enhanced voice file for a session using the production voice.

Frequently Asked Questions about Voice Synthesis

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate voice synthesis from SSML scripts for audio production?

You can automate voice synthesis from SSML scripts using Google Cloud Text-to-Speech to apply production defaults and psychoacoustic enhancements. This process automatically generates ready-to-mix, production-grade voice audio, significantly reducing manual production time.

What is the best way to generate long-form narration using Google Cloud TTS?

The best way to generate long-form narration with Google Cloud TTS is using automated chunking to process extended SSML scripts. This method applies a standardized production voice and preset parameters, ensuring consistent, high-quality audio output throughout the entire length of the script.

Do I need Google Cloud Text-to-Speech to produce hypnotic voice audio?

Yes, you need Google Cloud Text-to-Speech as the underlying dependency to produce hypnotic voice audio. It provides the en-US-Neural2-H production voice and neural processing required to apply psychoacoustic enhancements and generate high-quality, immersive session audio from SSML.

Can I standardize my voice branding output using SSML?

You can standardize voice branding output by applying a defined production voice, specifically en-US-Neural2-H, and preset parameters to your SSML scripts. This ensures consistent voice characteristics and audio quality across all generated production files.

Does this voice synthesis approach handle long hypnotic session scripts?

This voice synthesis approach handles long hypnotic session scripts by using automated chunking during the generation process. Chunking divides extended SSML scripts into manageable segments for Google Cloud TTS, ensuring production-grade audio output without exceeding processing limits.