sherpa-onnx-tts

Generate speech audio locally with sherpa-onnx text-to-speech models.

Updated Jan 29, 2026
One-click install
npx skills add https://github.com/douglasjs/clawdbot_plugin_tools --skill sherpa-onnx-tts-douglasjs
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: sherpa-onnx-tts
Source: https://github.com/douglasjs/clawdbot_plugin_tools/tree/main/sherpa-onnx-tts
Command: npx skills add https://github.com/douglasjs/clawdbot_plugin_tools --skill sherpa-onnx-tts-douglasjs

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) components.

What problem does it solve?

This Skill provides a local, offline text-to-speech (TTS) solution, eliminating the need for cloud-based services and ensuring privacy and accessibility.

Core Features & Use Cases

  • Local TTS: Generates speech audio directly on your machine without internet connectivity.
  • Customizable Voices: Supports various voice models for different speech characteristics.
  • Use Case: Convert written reports or documents into audiobooks for hands-free listening during commutes or while multitasking.

Quick Start

Use the sherpa-onnx-tts skill to convert the text "Hello from local TTS." into an audio file named tts.wav.

Frequently Asked Questions about sherpa-onnx-tts

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate text-to-speech audio offline without an internet connection?

Offline text-to-speech synthesis generates audio directly on your machine using the sherpa-onnx runtime and voice models, eliminating internet connectivity requirements and ensuring data privacy.

Does offline text-to-speech require specific runtime and model directories to be configured?

Yes, offline text-to-speech requires specific sherpa-onnx runtime and voice model directories to be configured on your device to facilitate on-device audio generation.

What is the best way to ensure privacy and reduced latency during audio generation?

Local offline text-to-speech synthesis ensures privacy and reduces latency by generating speech audio directly on your machine using sherpa-onnx, bypassing cloud-based service transmissions.

Can I customize speech characteristics using different voice models in on-device text-to-speech?

Yes, on-device text-to-speech supports various configurable voice models via the sherpa-onnx runtime, allowing you to customize speech characteristics for local audio generation.