voice-clone · Voice Cloning Dubbing

voice-clone

Clone your own voice from an uploaded sample, then synthesize talking-head narration, voiceovers, or sales copy in that voice. Runs on cloud providers (Alibaba CosyVoice, MiniMax, Fish Audio and more) — no local GPU needed.

Two steps: register the voice (enroll → voice_id), then synthesize any text with it. Bring your own provider API key. Compliance line: only clone your own voice or one you are explicitly authorized to use.

Example invocation: "Dub this talking-head script in my own voice."

Full brief

Positioning

voice-clone is the dubbing tool for "my voice": clone a personal voiceprint first, then synthesize any copy with it. It serves personalized-voice needs; for ready-made public voices use tts-voiceover instead.

Core capabilities

Workflow

  1. Pick a provider, configure its API key in .env, run check to verify
  2. Enroll: upload the voice sample → voice_id
  3. Synthesize: voice_id + text + speed → outputs/<topic>/
  4. (Optional) audio-mix for BGM, auto-subtitle for captions

Inputs & outputs

Input Required Notes
Voice sample Required for cloning Your own clean, noise-free voice (10s–1min per provider)
Text Required for synthesis The copy to speak in the cloned voice
provider Yes dashscope / minimax / fish-audio / openai-compatible / gemini

Output: synthesized speech audio (mp3/wav) under outputs/<topic>/.

Boundaries with adjacent skills

Skill Lane
voice-clone (this) Clone a personal voice, then synthesize; needs cloud keys
tts-voiceover Ready-made public voices (edge-tts); no key needed
ai-music AI-generated music/BGM
audio-mix Mix synth output with BGM

Fit

Before you start


voice-clone is part of the Aiglade Skill library. Invoke it from the Aiglade chat box in plain language.

← All 155 skills