Audio Toolbox - Voice Changer & Vocal Isolation
AI audio tools: change the voice while keeping the performance, strip background noise and music from recordings, and align existing audio to a script for exact subtitle timing.
Last updated: 2026-08-23
Three tools for audio you already have
Everything else in the studio generates from scratch; these three work on existing recordings.
Voice changer: keep the performance, swap the voice
Upload a recording, pick a target voice — pacing, pauses and emotion survive intact; only the timbre changes. Great for: performing a line yourself to control the delivery, then swapping in the character's voice; or fanning one dry read out into several character versions.
Real before/after · the same read, before (Sarah) → after (Callum):
Before:
After (identical performance, different voice):
Voice isolator: pull clean speech out
Strip background noise and music, keep the voice. Two frequent uses:
- Clean cloning samples: noisy recording? Isolate first, then clone — quality improves noticeably
- Rescue field audio: street interviews and event recordings become usable again
Forced alignment: recording + script = exact timing
Already have both the audio and the script, just missing timestamps? Forced alignment returns each word's precise start and end. Real output from our English homepage sample:
The [0.08s→0.14s] rain [0.16s→0.34s] had [0.36s→0.50s]
stopped, [0.52s→0.80s] ...
Versus speech to text: transcription guesses what was said; alignment already knows, and marks when — faster, more precise, zero typos when you have the script.
Where to find them
All three live in Voice Studio → Audio Tools: upload a file (≤200MB), the credit price sits on the button, failures refund automatically. Alignment results download as SRT.