Record or Prep the Cleanest Take You Can#
No chain fixes a bad recording. Before I touch a single device, I make sure the source material is worth mixing. I record 15–20 cm from the mic with a pop filter, keep input peaks around -12 to -6 dBFS, and track in the quietest, softest-furnished room available. A closet full of clothes beats an empty tiled bathroom every single time — my neighbors have confirmed this.
When I'm working with a vocal someone else recorded, I do the prep work instead: cut silence and breaths I don't want, consolidate clips, fix clipped words with a punch-in or a double, and fade clip edges so nothing clicks. Ten minutes of editing here saves an hour of trying to EQ a problem that is actually a mouth noise.
I set the vocal track fader so it sits a few dB above the instrumental before processing. Gain staging into the chain matters more than any single plugin setting — I learned that one the hard way.
Subtractive EQ Before Anything Else#
The first device in my chain is EQ Eight, and its job is removal, not enhancement. Everything I cut here means the compressor downstream reacts to the voice, not to rumble and mud.
- High-pass: I run 80–120 Hz, 12–24 dB/oct. Vocal fundamentals rarely live below 100 Hz; everything under it is mic handling noise, room rumble and the neighbor's washing machine.
- Low-mid cut: I sweep 200–500 Hz and pull down boxiness, usually 2–4 dB with a medium Q.
- Harshness notch: If the vocal bites, I look between 2–5 kHz and notch 2–3 dB. Sweep with a boosted narrow band to find the ugly spot, then cut it.
Resist the urge to boost yet — I still catch myself reaching for the top end too early. Additive EQ sounds better after dynamics, when you're shaping a controlled signal instead of a moving target.
"Cut first, boost second. The vocal you want is usually hiding under the vocal you recorded." — Monakai
Compression That Controls Without Crushing#
Ableton's stock Compressor handles vocals well if you stop asking one instance to do all the work. The classic move — and the one I use on nearly every vocal — is two gentle stages instead of one aggressive one.
Stage one, on the Compressor: I set a 3:1 ratio, attack around 10 ms so consonants keep their edge, release 80–120 ms (auto release also works), and a threshold that gives me 3–5 dB of gain reduction on the loud phrases. This levels the performance without strangling it.
Stage two, on the Glue Compressor: 2:1 ratio, slow attack (10–30 ms), fast or 0.2 s release, catching only 1–2 dB on peaks. Glue's soft knee and program-dependent behavior act like a polite finishing touch — it glues the vocal to itself after the first stage did the heavy lifting. I match output gain so my bypass comparisons stay honest.
"One compressor doing 6 dB sounds like a compressor. Two doing 3 dB sound like a singer who can sing." — Monakai
If you'd rather start from a tuned chain than build one from scratch, Fire Vocal Presets ships ready-made vocal chains with sensible staging already dialed in. I built it because I kept rebuilding the same chain session after session — it's a solid reference even if you end up tweaking every knob.
De-Essing and Resonance Control#
Sibilance lives roughly between 4 and 8 kHz depending on the voice, and compression tends to make it worse — the compressor clamps the body of the word and the "s" pokes through on release. That's why I de-ess after compression in this chain.
Using the Multiband Dynamics workaround from the tip above, I aim for 3–6 dB of reduction only while sibilance is active. If every "s" disappears entirely, you've gone too far; a vocal with no sibilance sounds like it was recorded underwater.
While I'm here, I listen for room resonances — those narrow, ringing frequencies that appear on certain notes, often between 500 Hz and 2 kHz in untreated rooms. I catch them with one or two narrow EQ Eight cuts of 2–4 dB, or automate the cut only on the offending word. The goal is a vocal that stays smooth from the quietest verse to the belted chorus.
Space: Delay and Reverb Sends#
Put space effects on return tracks, not on the vocal channel. Sends keep the dry signal intact and give you a separate fader to ride the effect level.
A practical setup: Return A holds a short, bright reverb for presence — decay around 1.2–1.8 s, pre-delay 20–40 ms so the dry transient stays in front, and high-pass the return at 250 Hz to keep low-end mud out. Vocal Verb was built for exactly this slot — a reverb tuned for vocals, so the decay and diffusion already sit where a voice wants them instead of where a hall preset thinks they should be.
Return B holds Ableton's Echo for vocal throws: dotted eighth or quarter-note delay, 30–40% feedback, and a low-pass on the repeats so each echo sits behind the dry vocal. Automate the send so the delay only fires on the last word of a phrase. A constant slapback on every word is how demos sound like demos.
One more trick: sidechain the reverb return to the dry vocal with a Compressor, 2–3 dB of ducking, fast attack, medium release. The reverb ducks while the singer sings and blooms into the gaps — space without wash.
Automation Is the Last 20 Percent#
A static vocal mix sounds finished for about four bars. Real mixes move. Once the chain is set, press A to open automation lanes and do a riding pass on the vocal fader: push ad-libs and doubles down 2–4 dB, lift word endings that get swallowed by the beat, and pull the whole verse down half a dB if the chorus needs somewhere to go.
Automate sends, too. Verses usually want less reverb than choruses, and delay throws belong on specific words, not everywhere. If a line still will not sit after fader rides, automate EQ Eight's output gain or a single band instead of adding another plugin — at this stage, subtraction and level moves beat more processing.
This is also the moment to sanity-check against reference tracks. Solo is a liar; make every final decision with the full beat playing.
Try It Yourself in Ableton#
Build the whole chain from this article in one session:
- Drop EQ Eight on the vocal track. High-pass at 100 Hz (24 dB/oct), then sweep 200–500 Hz and cut the boxiest spot by 3 dB with a Q around 1.5.
- Add the stock Compressor: ratio 3:1, attack 10 ms, release 100 ms, and set the threshold for 3–5 dB of gain reduction on the loudest phrase. Match makeup gain by ear with bypass toggling.
- Add Glue Compressor after it: ratio 2:1, attack 10 ms, release 0.2 s, catching only 1–2 dB on peaks.
- Build the stock de-esser: insert Multiband Dynamics, set the high-band crossover to 5.5 kHz, and compress only the top band — ratio 3:1, fast attack and release, 3–6 dB of reduction on "s" sounds.
- Create Return A with Vocal Verb (decay ~1.5 s, pre-delay 30 ms) and high-pass the return at 250 Hz. Send the vocal at around -12 dB to start.
- Create Return B with Echo: dotted eighth notes, 35% feedback, low-pass the repeats at 5 kHz. Automate the send so only phrase-ending words throw into the delay.
- Do one full fader-riding pass with automation, then compare the whole chain against a reference vocal at matched loudness. Adjust the first EQ and nothing else if something feels off.
If you want to compare your hand-built chain against a professionally staged one, Fire Vocal Presets gives you ready-made vocal chains you can A/B against your own — the fastest way to hear what your settings are missing.