Audio Toolbox - Voice Changer & Vocal Isolation

AI audio tools: change the voice while keeping the performance, strip background noise and music from recordings, and align existing audio to a script for exact subtitle timing.

Last updated: 2026-08-23

Three tools for audio you already have

Everything else in the studio generates from scratch; these three work on existing recordings.

Voice changer: keep the performance, swap the voice

Upload a recording, pick a target voice — pacing, pauses and emotion survive intact; only the timbre changes. Great for: performing a line yourself to control the delivery, then swapping in the character's voice; or fanning one dry read out into several character versions.

Real before/after · the same read, before (Sarah) → after (Callum):

Before:

After (identical performance, different voice):

Voice isolator: pull clean speech out

Strip background noise and music, keep the voice. Two frequent uses:

  • Clean cloning samples: noisy recording? Isolate first, then clone — quality improves noticeably
  • Rescue field audio: street interviews and event recordings become usable again

Forced alignment: recording + script = exact timing

Already have both the audio and the script, just missing timestamps? Forced alignment returns each word's precise start and end. Real output from our English homepage sample:

The [0.08s→0.14s]  rain [0.16s→0.34s]  had [0.36s→0.50s]
stopped, [0.52s→0.80s] ...

Versus speech to text: transcription guesses what was said; alignment already knows, and marks when — faster, more precise, zero typos when you have the script.

Where to find them

All three live in Voice Studio → Audio Tools: upload a file (≤200MB), the credit price sits on the button, failures refund automatically. Alignment results download as SRT.

Process my first file →