tts-voiceover · Text-to-Speech Voiceover

tts-voiceover

Turns scripts and copy into AI voiceover — spoken delivery, narration, read-aloud audio — with synchronized sentence-level SRT subtitles.

Closed-source cloud TTS (expressive, near-human) is the default when configured; without keys it falls back to edge-tts (flatter, mechanical). The voice track can be mixed with BGM or dropped into video as narration.

Example invocation: "Voice this script with a female narrator and generate a subtitle file alongside."

Full brief

Positioning

tts-voiceover is a text-to-speech dubbing tool: scripts and copy in, AI voiceover out (spoken delivery, narration, read-aloud). It does human voice — never background music.

Core capabilities

Workflow

  1. Optional: pick a voice with voices (common Chinese voices with notes; full list needs internet)
  2. speak to synthesize (optional synced SRT)
  3. Optional post-processing: mix / normalize / add to video

Inputs & outputs

Input Notes
text / file Required — text to voice or path to a text file (long text via --file)
voice Optional — voice id (default Xiaoxiao)
rate / volume / pitch Optional — speed / volume / pitch tweaks
output Optional — defaults to outputs/<topic>/{name}.mp3

Output: voiceover audio (mp3 / wav / m4a) with optional synced SRT, under outputs/<topic>/.

Boundaries with adjacent skills

Fit

Before you start


tts-voiceover is part of the Aiglade Skill library. Invoke it from the Aiglade chat box in plain language.

← All 155 skills