Free course

Complete Vocal Mixing & Mastering Course

Updated September 15, 2026 · PhiloStudios Team · ~25 min read · 10 modules

This course covers the whole path from a freshly recorded vocal to a final file ready to upload to Spotify or YouTube: session prep, dynamics, EQ, space, saturation, and a complete mastering chain. It doesn't matter whether you mix with PhiloStudios or any other tool — the principles are the same.

1Session prep and gain staging

Before inserting a single plugin, get your session organized: name your tracks, group lead vocals and backing vocals into a bus, and check your input level. Gain staging (calibrating signal levels at each stage) is the foundation of a good mix: if the vocal comes in too quiet, you'll end up boosting background noise along with it later; if it comes in too hot, you'll clip plugins that weren't designed for that.

A simple reference: aim for your loudest vocal peak to sit somewhere between -12 and -6 dBFS before any processing, leaving headroom for whatever plugins come next.

2Editing and cleanup

Before processing, edit the take: trim silences with background noise, even out overly loud breaths (turn them down a few dB rather than removing them entirely — a vocal with zero breaths sounds artificial), and comp together the best parts of multiple takes if you recorded more than one.

Also check for clicks, mouth noises or the odd mic pop; sometimes it's worth fixing these by hand rather than expecting a plugin to solve them automatically.

3Gate: removing background noise

A gate automatically cuts or attenuates the signal when it falls below a threshold, letting the vocal through only once it crosses that level. It's the first dynamics processor in the chain because it works best on the rawest possible signal — apply it after compression instead, and the compressor will have already lifted that background noise along with the vocal, making the gate far less effective.

Key parameters

  • Threshold: the level below which the gate starts to close. Set it by ear, based on where the background noise sits relative to the vocal.
  • Attack: how fast the gate opens. Too slow can clip the start of words.
  • Release: how fast it closes. Too fast can sound choppy and unnatural between phrases.

4Compression: evening out dynamics

The human voice varies a lot in volume between words and phrases. A compressor reduces that gap so the vocal stays present and consistent throughout the song. A good starting point for vocals is a 3:1 to 4:1 ratio, medium attack (5-15 ms), and auto or medium release, adjusted until gain reduction sits around 3-6 dB on the loudest parts.

Serial vs. parallel compression

Instead of one aggressive stage, many engineers use two or three gentle compressors in series (each reducing a little) for a more transparent, natural result than one hard compression. Parallel compression is another technique: you blend the original signal with a heavily compressed copy, gaining density and presence without losing the natural dynamics of the performance.

5Subtractive and additive EQ

EQ decides where the vocal "lives" in the frequency spectrum, and where it sits in the mix relative to the instruments. It helps to think of it in two passes:

Subtractive EQ (first)

Applied early to remove problem frequencies before compressing: specific resonances, excess low end, or muddy zones.

Additive EQ (later)

Applied once dynamics are already under control, to add brightness and presence.

RangeWhat to do there
80-120 HzLow cut: removes rumble and mic proximity noise.
200-400 HzToo much energy here sounds muddy; a gentle cut cleans up the mix.
2-5 kHzPresence and intelligibility zone; a subtle boost helps the vocal cut through the instrumental.
10 kHz and up"Air." A gentle shelf boost adds brightness without adding harshness.

6De-essing: controlling sibilance

The consonants "s," "sh" and "ch" generate concentrated energy typically between 4 and 9 kHz that, after compressing and EQ'ing (especially if you boosted presence), can become harsh. A de-esser is a compressor specialized on that band: it automatically reduces those frequencies only when sibilance appears, without affecting the rest of the vocal. It works well after additive EQ, since that boost is often exactly what makes sibilance more noticeable in the first place.

7Space: reverb, delay and stereo width

Reverb and delay place the vocal inside an acoustic space and add depth. Use them as sends rather than inserting them directly on the track, so you can precisely control the wet/dry balance. They come after dynamics: compressing a signal that already has reverb on it also compresses the reverb tail, which can sound unnatural or "pumped."

A short tempo-synced slap delay adds thickness and presence without sounding like an obvious echo. A plate or room reverb adds cohesion with the rest of the mix without pushing the vocal too far back. For backing vocals or doubles, some stereo width helps separate them from the lead, which you'll usually want to keep centered and mono-compatible.

8Saturation and color

Gentle saturation adds harmonics that make the vocal feel more present and "glued" to the mix — especially useful in urban and pop genres where a vocal with character and density is the goal. Used in moderation, it helps the vocal cut through busy instrumentals without needing to raise the volume. Overused, it can sound dirty or distorted — think of it as a seasoning, not the main dish.

9Mastering: the difference and the typical chain

Mixing works the balance between all the elements of a song (vocals, drums, bass, instruments). Mastering is the final step: it takes that finished mix (in stereo, a single file) and prepares it for distribution — making sure it sounds good on any playback system and has the right perceived loudness compared to other commercial tracks.

Typical mastering chain

  1. Corrective EQ: gentle, broad adjustments (not surgical like in mixing) to balance the overall spectrum.
  2. Multiband compression: controls dynamics independently across frequency bands — useful for taming lows or highs that stick out without affecting everything else.
  3. Saturation or harmonic exciter (optional): a final touch of color and density.
  4. Limiter: raises overall volume while controlling peaks, to reach the target loudness without distortion.

A golden rule: if you need drastic EQ or dynamics changes at the mastering stage for a song to sound right, the problem is probably in the mix, not the mastering. Mastering polishes; it doesn't fix an unbalanced mix.

10Loudness (LUFS) for streaming and final checklist

Streaming platforms automatically normalize volume, so mastering "as loud as possible" no longer gives you an edge — if anything, it can sound worse if the limiter had to work too hard. Each platform targets a different loudness level, measured in integrated LUFS:

PlatformApproximate target
Spotify≈ -14 LUFS
YouTube≈ -14 LUFS
Apple Music≈ -16 LUFS

You don't need to chase the exact number — if your master ends up a bit louder, the platform will turn it down anyway. What matters is not over-compressing to chase loudness you won't actually keep.

Final checklist before exporting

  • Listen to the mix on different systems: speakers, headphones, your phone's built-in speaker.
  • Check for clipping (peaks above 0 dBFS) in the final master.
  • Compare your mix to a commercial reference in the same genre, at matched volume.
  • Rest your ears and listen again the next day before calling the master final.
Apply all of this in one plugin

Gate, parametric EQ, serial compressors, de-esser, reverb, delay and saturation: PhiloStudios brings this course's entire vocal mixing chain into a single VST3, with factory presets for different genres. Download the free demo and try it on your own mix.

Listen: before/after →