Vocal chain: the plugin order for a clear, upfront vocal
4 min read · Updated
A vocal that sits in front of the mix without hurting the ear is mostly a matter of order: each process prepares the next one. Here is the most common studio vocal chain, step by step, with starting settings that work on most voices, in rap, pop or singer-songwriter music.
In short
- Clean the take: background noise, loud breaths, bleed.
- Corrective EQ: cut useless low end and resonances.
- Compression: steady the level so every word is understood.
- De-esser: tame the "s" and "sh" sounds compression brings out.
- Color: light saturation and presence EQ.
- Space: reverb and delay on aux sends, never too much.
First things first: a clean take
No plugin replaces a good take. Place the mic 15 to 20 cm away with a pop filter, and aim for peaks around -10 dBFS when recording: you keep headroom without raising the noise floor.
If the room or the air conditioning can be heard, clean first, before anything else: a denoiser placed after a compressor has to fight noise the compression already pushed up. For spoken word recorded in an untreated room, NAUPLY Clarea chains local neural denoising with breath and plosive reduction; on singing, keep it light and check by ear. To simply cut noise between phrases, a noise gate like Gatea is enough.
Corrective EQ: remove before you add
Start with a high-pass filter between 70 and 120 Hz depending on the voice: it removes rumble, mic-stand bumps and low end that would clutter the mix without adding anything to the vocal.
Then look for resonances: sweep a narrow, boosted bell slowly, then cut 2 to 4 dB where the sound turns nasal (often 300 to 600 Hz) or harsh (2 to 4 kHz). An EQ that shows the live spectrum, like Eqora, lets you spot these areas by eye and confirm by ear.
Compression: every word at the same level
Compression reduces the gap between loud and quiet passages. Starting point: a 3:1 to 4:1 ratio, a medium attack (10 to 30 ms) to let consonants through, a 50 to 150 ms release, and a threshold that gives 3 to 6 dB of gain reduction on loud passages.
If the vocal is still uneven, two gentle compressors in series beat one pushed hard: the first catches peaks, the second smooths the whole. Parallel mix, offered by Compra, keeps the voice natural while adding density.
De-esser: sibilance under control
Compression and presence EQ bring out "s", "sh" and "z" sounds. A de-esser only turns the sound down when sibilance crosses a threshold in a specific frequency band, usually between 5 and 9 kHz for a voice.
Listen to the monitored band (Deessa's LISTEN function) to lock it exactly on the sibilance, then reduce until it stops stinging. Too much reduction causes a lisp: if the voice loses its "s" sounds, you went too far.
Color and presence
Light saturation (tube or tape) adds harmonics that make the voice denser and easier to hear on small speakers, without raising the level. Blend it in parallel: a few percent is often enough. Tubra offers both characters, with a perfectly aligned dry / saturated mix.
Finish with gentle presence EQ: a broad +1 to +3 dB shelf above 8 to 10 kHz for air, possibly a slight bump around 3 kHz for intelligibility. Always compare at equal loudness: louder always sounds better.
Placing the vocal in space
Reverb and delay work best on aux sends: the dry vocal stays upfront and the space is set with the send level. A 20 to 60 ms pre-delay separates the voice from its tail and keeps lyrics readable; also cut the lows and extreme highs of the return.
For a modern vocal, a tempo-synced delay (eighth or dotted quarter) often works better than a big reverb; automate "throws" at the end of phrases to keep the ear engaged. At NAUPLY, that is the job of Lontana Reverb (plate, room, hall) and Lontana Delay (digital, analog or modulated).
Going faster: a chain set by analysis
Building this chain by hand is still the best way to understand your voice. To go faster, Vocora analyzes the vocal passing through it (pitch, sibilance, dynamics) and sets the whole chain at once: high-pass, corrective EQ, de-esser, compressor, presence, saturation and space. Every module stays adjustable: it is a starting point, not a black box.
Frequently asked questions
- Should the de-esser go before or after the compressor?
- Both work. After the compressor, it catches the sibilance compression brought out; before, it keeps sibilance from triggering the compressor. Most common: after compression, plus a second light de-esser at the end of the chain if presence EQ wakes it up again.
- Which reverb for a rap vocal?
- A plate or a short room (0.6 to 1.2 s), with pre-delay and a low level, often paired with a synced delay automated at the end of phrases. The goal is depth without pushing the vocal back.
- Is the vocal chain the same for spoken word?
- Same principle, with more cleanup (noise, breaths, plosives), steadier compression for podcasts and almost no reverb.
NAUPLY news
Plugin releases, updates and offers. A few emails a year, one-click unsubscribe.