Desktop · 0:53 video
Preview a change to the spoken words
Edit a short line, hear a local voice preview, then explicitly accept or undo the audio replacement.
Before you start: A transcript, the optional Pocket TTS pack, and permission to use the speaker’s voice.
Download videoDownload captions
Step by step
Select the words to replace
Select a short passage in Transcript. Open Speakers and wording → Change spoken words…. This is separate from Correct word, which edits transcript text only.
Set the line and reference
Enter New spoken wording. Listen to a clear 2–10 second reference from the same speaker, and confirm that you have permission to use their voice.
Generate and listen
Choose Generate voice preview, then Listen to voice preview. The preview preserves the selected passage’s duration and includes the edit fades. Generate again after changing the wording.
Accept only a useful result
Accept audio replacement keeps the generated audio and matching words as an undoable edit. Ctrl + Z restores the original audio and transcript together.
Listen to the generated voice preview. Open full size
- The revised spoken line is separate from the original wording.
- Listen to a clear reference with permission.
- Listen to the generated preview.
- Accept only after reviewing the result.
Read the video instructions
- This example uses a synthetic voice. Use a real speaker’s voice only with their permission.
- 1. Select the spoken passage. Open Speakers and wording → Change spoken words…
- 2. Edit the line. Check the voice reference and confirm permission to use that voice.
- 3. Generate voice preview. Generation runs locally on the CPU.
- Processing runs locally. Wait for the result before continuing.
- The result is ready to review.
- Listen to the preview before accepting. The passage keeps its length and edit fades.
- 4. Accept audio replacement keeps the preview. Ctrl + Z restores the original audio and words.
- Keep only a preview you are happy with. Correct word remains the simpler text-only correction tool.
Audio credits
LibriSpeech excerpts 1688-142285-0007 and 1272-128104-0000, collected by Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur from LibriVox, are used under CC BY 4.0. Excerpts have been trimmed, resampled, padded, or processed. The voice-replacement example uses an original synthetic voice.