In 1939, engineers at Warner Bros. built the first de-esser. The sharp “s” sounds in an actor’s voice were overloading the optical film strips that carried sound, so they made a system to tame that energy automatically.
Nearly ninety years later, engineers are still fighting the exact same enemy.

✨ Summary
Sibilance is the harsh “s” and “sh” energy in a voice, and it hits hardest on acapellas, where no instruments mask it and mastering only amplifies it. This guide shows you how to fix it: sweep to find the exact frequency, set a de-esser properly, ride the worst syllables by hand, and use dynamic EQ for stubborn spots, all without creating a lisp.
Here’s what most people don’t realize: sibilance isn’t a flaw in your singer or your mic. It’s baked into human speech. Every voice has it. The only question is whether you control it or let it control your record.
And on an acapella, that question gets loud. With no instruments sharing the high end, a harsh “s” has nowhere to hide, and the same sibilance that slips by in a full mix can turn a solo vocal into something painful through earbuds.
This guide breaks down what sibilance is, why a bare vocal makes it worse, and how to fix it properly without giving your singer a lisp.
These might be your queries if you’re recording raw acapella vocals:
- Every ‘s’ and ‘t’ felt like it was cutting through my ears
- The high frequencies were very sharp, almost piercing
- “unbearable on headphones” when I listen to
- “painful to listen to” even on speakers
If these queries feel like your own and you’re constantly looking for solutions, I welcome you to this blog.
First, we’ll start with an understanding of acapella from a technical perspective.
What Sibilance Actually Is
Sibilance is the sharp, hissy burst of high-frequency energy that comes out on certain consonants. The obvious one is “s,” but it’s also “sh,” “ch,” “t,” “z,” and “j.” Say “she sells seashells” out loud and you can feel exactly where it lives.
That energy sits high in the frequency spectrum. Depending on the voice, sibilance falls somewhere between roughly 2 kHz and 10 kHz, with most of the trouble concentrated in the 5 kHz to 9 kHz range. Where exactly depends on the singer. A deeper voice tends to sit lower, a brighter voice higher.
It’s not a defect. It’s how speech works. The problem is only that some voices, some microphones, and some rooms push that energy past the point of comfort. And once you start mastering for loudness, every bit of it gets amplified.
Sibilance is always there. Mastering just decides how loud it gets.
If the frequency ranges here feel unfamiliar, this beginner’s guide to music frequencies is worth a quick read before you go further. The rest of this gets easier once those numbers mean something to you.
How does a cappella sound?
An acapella is the human voice completely on its own — no instruments, no backing, just breath, pitch, and words carrying the entire song. It sounds intimate and exposed, every detail laid bare, which is exactly why flaws like harsh “s” sounds have nowhere to hide.
Why Acapella Makes It Worse
In a full mix, sibilance has competition. Cymbals, guitars, synths, and hi-hats all live in that same high range, so they blur and mask the “s” sounds. Your ear has a dozen things to focus on, and the sibilance blends into the texture.
Strip the instruments away, and that cover disappears completely.
Now the voice owns the entire high end by itself. There’s nothing sharing that frequency space, nothing distracting the listener, nothing softening the edge. Every “s” stands alone, fully exposed, with the listener’s attention pointed straight at it.
Then mastering makes it worse. To bring a solo vocal up to a competitive streaming loudness, you raise the overall level, and the sibilance rises right along with everything else. The harder you push for volume, the sharper those consonants get.
This is why sibilance has to be handled early. Fix it before you chase loudness, because loudness makes it louder.
Step 1: Find the Exact Frequency
You can’t fix what you can’t locate. Before reaching for any tool, find the precise frequency where your singer’s sibilance lives, because it’s different for every voice.
- The technique is simple. Open an EQ, create a narrow bell, and boost it hard, somewhere around +10 to +15 dB.
Then slowly sweep that boosted band across the high frequencies, from about 4 kHz upward, while the vocal plays.
When you hit the sibilant frequency, it’ll jump out at you. The “s” sounds will become piercing, almost unbearable. That’s your target.
Make a note of where it peaks. Most voices land between 5 kHz and 8 kHz, but yours might sit a little above or below. Once you know the exact spot, every tool you use afterward becomes far more accurate, because you’re aiming instead of guessing.
Step 2: Use a De-Esser the Right Way
The de-esser is the dedicated tool for this job, and it’s been refined for decades, from the first standalone units in the 1970s to the studio standards that followed.
It works by watching the sibilant frequency range and turning it down automatically, but only in the instant a harsh “s” spikes through. The rest of the vocal stays untouched.
To set it well:
- Point it at the right frequency. Use the target you found in Step 1. Most de-essers let you set the band, so center it where the sibilance actually peaks instead of leaving it on a generic default.
- Set the threshold by ear. Lower the threshold until the harsh peaks get caught and softened, then stop. You want it reacting to the worst “s” sounds, not every word.
- Listen for over-correction. Push too hard and you’ll hear it. The “s” sounds start to vanish entirely, and the singer develops a soft, lisping quality. That’s the sign you’ve gone too far. Back it off.
The goal isn’t to remove the sibilance. It’s to make it stop demanding attention. Done right, a de-esser is invisible. You simply stop noticing the “s” sounds.
Step 3: Ride the Worst Offenders by Hand
Here’s the move that separates a clean master from a great one.
A de-esser treats every sibilant sound the same way, based on a threshold. But real performances aren’t uniform. Usually only a handful of words have genuinely harsh “s” sounds, while the rest are fine. Crank the de-esser hard enough to fix those few bad ones, and you dull the entire vocal in the process.
So fix them individually instead.
Go through the track and find the specific syllables that spike. Then use clip gain or volume automation to pull down just those moments, by a decibel or two, by hand. It’s slower, but it’s surgical. You’re correcting the three or four problem words instead of taxing every “s” in the song.
This lets you run the de-esser much more gently, because it no longer has to fight the worst offenders alone. The two approaches work together: manual rides handle the spikes, the de-esser smooths the rest.
The cleanest results almost always come from doing less to the whole and more to the few.
Step 4: Reach for Dynamic EQ When You Need Precision
Sometimes a de-esser is too blunt. It clamps down on a whole band when you really only want to touch one narrow sliver of it. That’s where dynamic EQ comes in.
A dynamic EQ lets you set a tight, specific band, exactly the frequency you found in Step 1, and have it reduce the level only when that frequency gets too loud. It’s like a de-esser with a sniper scope: narrower, more targeted, and gentler on everything around it.
Use it when the sibilance is concentrated in one very specific frequency and a standard de-esser is grabbing too much of the surrounding brightness. You keep the air and clarity of the voice while still taming the harsh spike.
For most acapellas, a good de-esser plus a little manual riding handles the job. But when a particular voice has one stubborn, pinpoint problem frequency, dynamic EQ is the cleanest tool for it.
The Mistakes That Make It Worse
Most sibilance disasters come from the same handful of errors.
- Over-de-essing. The big one. Push a de-esser too hard and you suck the life out of the consonants, leaving a dull, lisping vocal that sounds processed and lifeless. A little harshness is far better than an obvious lisp.
- Fixing it after you’ve maximized loudness. If you push the master loud first and de-ess second, you’re fighting amplified sibilance. Handle the “s” sounds early in the chain, before the heavy loudness processing.
- Boosting the air band too hard. Everyone loves that bright, expensive “air” up top. But boost the high end aggressively and you’ll drag the sibilance right back up with it. Brightness and de-essing pull against each other, so balance them.
- Judging it on the wrong speakers. Sibilance reads completely differently on cheap earbuds than on studio monitors, and what sounds smooth in your room can be brutal on a phone. Honest monitoring matters, and how your speakers are set up changes what you’re able to hear in the first place. Always check the final result on a few different devices.
Avoid those four, and you’re most of the way there.
Where Remasterify Fits
Fixing sibilance by hand is precise work. Sweep for the frequency, dial in a de-esser without overdoing it, ride the worst syllables individually, maybe add a dynamic EQ for the stubborn spots, then re-check the whole thing once loudness is applied. On a bare vocal, where there’s nothing to hide behind, the margin for error is thin.
This is the kind of balancing act Remasterify is built to handle. Instead of slapping a fixed de-esser across everything, it analyses your specific vocal, finds where the harsh energy actually sits, and controls it in proportion, then sets the track to a proper streaming loudness without letting the sibilance climb back up in the process.
It happens in seconds, with manual controls there if you want to push it further yourself.
You sang it clean.
Don’t let one sharp “s” be the thing people remember.
Master your acapella so every word lands smoothly on every speaker.
FAQ
Sibilance generally sits between 2 kHz and 10 kHz, with most of the harshness concentrated around 5 kHz to 9 kHz. The exact frequency varies from voice to voice, which is why sweeping to find your singer’s specific spot is the best first step.
In a full mix, instruments like cymbals and guitars share the high frequencies and mask the “s” sounds. On an acapella there’s nothing covering that range, so every sibilant consonant is fully exposed, and mastering for loudness amplifies it further.
A de-esser is a dedicated tool that ducks a sibilant band when it spikes. A dynamic EQ does something similar but lets you target a narrower, more specific frequency, making it more precise when the harshness is concentrated in one stubborn spot.
That’s over-de-essing. When you reduce the sibilant frequencies too aggressively, the “s” and “sh” sounds lose their definition and the voice starts to lisp. Back the de-esser off until the consonants return, and handle the worst spikes by hand instead.
Before. Raising the overall loudness lifts the sibilance along with everything else, so it’s far easier to control the “s” sounds early in the chain, then apply your loudness processing, rather than trying to tame amplified sibilance at the end.
