multi-voice-dubbing · Multi-Voice Dubbing

multi-voice-dubbing

A multi-character dialogue dubbing skill: from a cast table and line-by-line dialogue, it assigns each character a distinct voice and emotional register, producing a multi-voice audio track plus speaker-labeled subtitles.

Per-line emotion tags feed the voice engine's true emotion channel to drive the performance; narrator/host gets a separate voice from all characters. Broadcast-grade "sounds human" results need a cloud voice provider — the free engine is draft-grade only.

Example invocation: "Dub this two-person interview — a steady male voice for the host, a lively female voice for the guest."

Full brief

Positioning

multi-voice-dubbing turns "multi-person dialogue / multi-character scripts" into multi-voice audio: each character gets a voice that fits their persona — never one voice for the whole piece. The resulting voice.mp3 feeds straight into video assembly; voice.srt is speaker-labeled, time-aligned subtitles.

Core capabilities

Workflow

  1. Casting (cast.json): assign each speaker a voice, verify no collisions
  2. Line-by-line dialogue (lines.json): split into ordered lines, tag each with an emotion
  3. Synthesize (multivoice.py dub): outputs voice.mp3 + voice.srt
  4. Into video: use as narration for the video assembler, or mix with BGM

Inputs & outputs

Input Notes
cast.json Required — cast table: each speaker → voice and engine config
lines.json Required — ordered lines [{speaker, text, emotion}]
Output path Optional — defaults to voice.mp3 + same-name srt

Output: multi-voice track voice.mp3 and speaker-labeled aligned subtitles voice.srt.

Boundaries with adjacent skills

Fit

Before you start


multi-voice-dubbing is part of the Aiglade Skill library. Invoke it from the Aiglade chat box in plain language.

← All 155 skills