Audio
Runs in your browser

Audio De-Esser

Reduces harsh S and SH sounds by attenuating a sibilance band when that band gets loud. Frequency, threshold and amount control how much is reduced. A strong setting can dull the voice, so compare the preview before you export MP3, WAV or M4A. This is not a full vocal cleanup. Processing stays in this browser.

Accepted formats:
MP3WAVM4AOGG

The de-esser turns down a bright band when that band gets loud. Frequency is in hertz. A strong amount can dull the voice. This is not a full vocal cleanup.

Upload a voice recording

The file stays in this browser.

Preset

What does Audio De-Esser do?

Audio De-Esser turns down a bright band when that band gets loud. The file is split into a dry path and a side path. The side path is a high-pass at the frequency you choose, followed by a low-pass at 12 kHz, so the detector is listening to sibilance rather than the whole voice. When that band crosses the threshold, a compressor reduces the dry voice. Amount is the ratio. Attack and release are how fast the reduction starts and lets go. The frequency control is hertz, from 2 kHz to 12 kHz. A light setting catches only the harshest peaks. A strong setting can dull the voice because it also reduces brightness you may have wanted. This is not a full vocal cleanup, and it will not remove noise, room tone or distortion. Compare the before and after players. If the source already peaks near full scale, lower the amount before you export.

How to use Audio De-Esser

  1. 1Upload a voice recording and look at the waveform.
  2. 2Start with Light or Voice, or set frequency, threshold and amount yourself.
  3. 3Export and compare the after player with the original.
  4. 4Cancel stops the local job. Strong settings can dull the voice.

Supported inputs and outputs

Inputs

Local audio the browser can decode, such as MP3, WAV, M4A or OGG

Outputs

MP3, WAV, M4A

Practical uses

  • Softening harsh S and SH sounds in a narration
  • Trying a lighter pass before a stronger one
  • Checking a voice recording against a waveform before export

Browser and media limitations

  • A strong amount can reduce brightness and make the voice sound dull.
  • This does not remove noise, echo, clipping or a whole mix problem.
  • The band is fixed. It does not find S sounds by recognizing speech.
  • The media engine downloads on first export.

How local processing works

Your media is processed locally in your browser and is not uploaded to Tubelexity for processing. The media engine downloads on first export. Page navigation and analytics still use network requests.

Sibilance, sidechain reduction and dulling

Sibilance is the bright energy in S and SH sounds, often from about 5 kHz upward. The side path isolates that band in hertz and uses it as the signal that tells the compressor when to turn the voice down. The rest of the voice is not filtered away.

If the amount is high, ordinary brightness is reduced along with the harsh peaks. That is overprocessing. Use the after player and back the amount off if the voice sounds muffled.

Frequently asked questions

Which frequencies are treated as sibilance?

The side path starts at the frequency you set, between 2 kHz and 12 kHz, and stops at 12 kHz. Speech S sounds often sit in that bright range.

Will a strong preset clean the voice completely?

No. Strong reduction can dull the voice. It does not remove noise or fix a clipped recording.

What do threshold and amount do?

Threshold is how loud the bright band must be before reduction starts. Amount is how hard that band turns the voice down.

Is the recording uploaded?

No. The file stays in this browser tab. The media engine downloads on first export.

Related creator tools