Skip to content
VocalCopyCat

Voiceover editing and mixing

De-Ess Voice-Over Without Making Consonants Sound Lispy

Reduce harsh voice-over sibilance with targeted de-essing, local repairs, and sentence-level checks that preserve intelligibility and the speaker's character.

De-essing reduces distracting sibilance while keeping words recognizable. The target is an overly sharp consonant, not every trace of “s,” “sh,” or breath. If the processed speaker develops a lisp or the whole voice becomes dull, the treatment is doing more than the problem requires.

Confirm that the issue is sibilance

Listen to several examples in the raw voice. Does the harshness appear mainly on sibilant consonants, or do loud vowels also sound crunchy? Distortion from overload and a low plosive thump need different solutions.

Use the plosives-versus-sibilance guide when the diagnosis is unclear. For a new recording, a better microphone angle or another clean take may reduce the need for processing.

Select a test passage that includes a troublesome “s” and several normal ones. A setting based only on the worst syllable can overprocess the rest of the voice.

Locate the troublesome range by listening

Open a copy in an external editor with an appropriate de-esser. Adobe describes its DeEsser effect as targeting sibilance and provides controls for the affected range and processing behavior.

Different speakers and recordings place their distracting energy differently. Do not choose a frequency solely from a generic chart. If the effect provides a way to audition what it detects, use that briefly to locate the problem, then return to normal listening.

Your aim is to hear mostly the unwanted consonant in the detector audition, not substantial vowels or the entire brightness of the voice. The final decision still belongs to the whole sentence.

Apply reduction only as strongly as needed

Start with restrained reduction and play through the test passage. Adjust until the sharp syllables stop demanding attention while ordinary consonants remain intact.

Compare this original phrase: “Save the session before closing the workspace.” The word endings should stay clear. If “session” begins to sound softened into a different articulation, reduce the treatment or narrow its application.

Do not use the largest possible reduction simply because the loudest “s” becomes smooth. Speech identity and intelligibility live partly in these small consonant details. A slightly bright but clear word can be preferable to a heavily altered one.

Repair exceptions locally

If one syllable remains unusually sharp, try a local level adjustment on a duplicate or nondestructive clip. Preserve its start and end so the word does not acquire a click or unnatural gap.

A single exceptional sound does not always justify stronger processing across an entire paragraph. Likewise, a broadly bright voice may benefit from modest EQ, while a dynamic sibilance problem needs event-specific control. Listen before deciding which job you are doing.

Check the effect after other processing too. Compression or a brightness boost can change how prominent consonants feel. There is no universal effect order that removes the need to review the complete chain.

Keep one untreated comparison clip beside the working version. If several rounds of adjustment have made the entire voice less distinct, return to that reference and rebuild the treatment around a smaller set of clearly troublesome sounds.

Review at normal speed and ordinary volume

Listen to the entire passage without looping the repaired syllable. Repetition can make a normal consonant seem more objectionable than it will be in context. Take a short listening break if you stop trusting the comparison.

Check both headphones and a simple speaker. Listen for lost plurals, weakened word endings, a new lisp, or a dull quality that was not present before. Match playback level when comparing versions so a quieter result is not automatically judged smoother.

For generated narration, choose a suitable candidate in the VocalCopyCat studio before investing in detailed repair. Sometimes a revised phrase or another generation is the simpler production choice. Download the selected audio, de-ess only where necessary in an external editor, and retain the untouched file alongside the final version.

Sources and further reading

Production advice and sample scripts are editorial guidance. Check the linked documentation for current platform requirements.

Ready to find your next voice?

Listen to the samples, try a short preview, then create speech with prepaid credits in your workspace.

Open your studio

Hear your words come to life

Choose a voice and try a short preview.

77 / 120 input characters
Continue with 2,000 welcome credits

Listen to sample voices

Hear examples before choosing a voice. Generated results can vary with the script and reference sample.

Looking for another voice?

Explore the library and listen to a sample before you create.

Morgan Freeman avatar

Morgan Freeman

Morgan Freeman voice sample0:00 --:--
Stephen Hawking avatar

Stephen Hawking

Stephen Hawking voice sample0:00 --:--
Christiano Ronaldo avatar

Christiano Ronaldo

Christiano Ronaldo voice sample0:00 --:--
Donald Trump avatar

Donald Trump

Donald Trump voice sample0:00 --:--
Kokoro avatar

Kokoro

Kokoro voice sample0:00 --:--
Disney XD Announcer avatar

Disney XD Announcer

Disney XD Announcer voice sample0:00 --:--
Cute Japanese Girl avatar

Cute Japanese Girl

Cute Japanese Girl voice sample0:00 --:--
Vin avatar

Vin

Vin voice sample0:00 --:--
Adam Stone avatar

Adam Stone

Adam Stone voice sample0:00 --:--

Share this article