Mixing

Mixing Vocals Over a Leased Beat: A Step-by-Step Workflow

Mixing Vocals Over a Leased Beat: A Step-by-Step Workflow

A practical workflow for fitting your vocal into a finished leased beat, with real starting settings for EQ, compression, de-essing, autotune, BPM-synced effects, ducking and loudness.

Leasing a beat gets you a finished, mastered instrumental. It also hands you a specific mixing problem: you're fitting a raw vocal into a track that was built, balanced and often limited before you ever pressed record. Unlike a producer mixing a session from scratch, you can't turn the snare down or pull the lead synth out of your vocal's way, at least not if you only have the 2-track. Your mix decisions have to work around the beat instead of rebuilding it.

This article is a practical, step-by-step workflow for mixing vocals over a leased beat, with real settings you can start from. The numbers are starting points, not rules. Every voice, microphone and beat is different, so let your ears make the final call. If you haven't recorded yet, start with our guide to recording vocals at home on a type beat. A clean recording makes every step below easier.

2-Track vs Trackouts: Know What You're Working With

The first question is what files your lease includes, because that decides how much control you have.

File typeWhat you getMixing controlTypical lease
MP3 2-trackOne stereo MP3 of the full beatLowest. Whole-beat EQ and ducking only, lossy sourceMP3 Lease
WAV 2-trackOne stereo uncompressed WAVLow to medium. Full-quality source, still one fileWAV Lease
Trackouts (stems)Separate WAVs for drums, 808, melody, FX, etc.High. Rebalance, sidechain individual parts, carve space preciselyTrackout / Unlimited Lease

With a 2-track, your tools are broad EQ on the whole beat, sidechain or dynamic EQ ducking, and careful vocal processing. With trackouts you can turn down the lead synth during verses, sidechain only the melody to the vocal, or mute the 808 for a drop. Compare tiers on our licensing page. If you're releasing a serious single, the Trackout Lease often pays for itself in mix quality.

Tip: Always mix from WAV, never MP3, when you have the choice. Every processing step exaggerates MP3 artifacts, especially in the hi-hats and reverb tails above 10 kHz.

Session Setup and Gain Staging

Set up the session

  1. Create a project at the exact BPM of the beat (it's listed with the beat or in the file name) and the same sample rate as the files, usually 44.1 or 48 kHz.
  2. Import the beat (or stems) and line them up at bar 1.
  3. Create tracks for lead vocal, doubles, ad-libs and harmonies, and route them to a vocal bus.
  4. Create two effect return tracks: reverb and delay.
  5. Route the beat to its own beat bus so you can process the instrumental as one unit.

Gain staging

Leased beats usually arrive already mastered and loud, often peaking near 0 dBFS. If you stack vocals on top, the master clips before you've started. Fix that first:

  • Turn the beat bus down by 6–10 dB so it peaks around -6 dBFS.
  • Clip-gain each raw vocal so it averages around -18 dBFS RMS and peaks around -10 to -6 dBFS. Plugins emulating analog gear behave best around this level.
  • Keep the master fader at 0 dB with no limiter until the final step.

Before any processing, set a rough balance with just faders. The lead vocal should sit clearly on top of the beat, and you should understand every word at a moderate listening volume.

Cleanup and Vocal EQ

Cleanup first

Edit before you process. Cut or fade breaths that are too loud (reduce them 6–10 dB instead of deleting them completely, so the take stays natural), remove clicks and mouth noise, and gate or manually clean the silence between phrases. Tightening doubles and ad-libs to the lead with manual edits or an alignment tool makes stacks sound professional instead of phasey.

Subtractive EQ

Start by removing problems:

  • High-pass filter at 80–120 Hz for most male voices and 100–150 Hz for most female voices, 12–18 dB/octave. On ad-libs and doubles, go higher (150–250 Hz).
  • Mud: a wide cut of 2–4 dB around 200–400 Hz if the voice sounds boxy or muddy.
  • Nasal or honky tones: a narrow cut of 2–5 dB somewhere in 800 Hz–1.5 kHz. Sweep a narrow boost to find it, then cut.
  • Harshness: a narrow or dynamic cut around 2.5–5 kHz if the voice gets piercing when the artist pushes.

Additive EQ

After compression (or on a second EQ later in the chain), add character:

  • Presence: a gentle 1–3 dB bell boost around 3–5 kHz helps intelligibility over bright trap melodies.
  • Air: a high shelf of 2–4 dB from 10–12 kHz up. Use a de-esser after it.
  • Body: if the voice is thin, 1–2 dB around 150–250 Hz, but check that it doesn't clash with the 808 harmonics.

Processing doubles, ad-libs and harmonies

Supporting vocals shouldn't compete with the lead. Treat them as a separate layer with its own rules:

  • Doubles: sit 6–10 dB under the lead, high-passed higher (150–250 Hz) and slightly darker, with a low-pass around 10–12 kHz. Pan a pair of doubles 30–60% left and right for width, and keep the lead dead center.
  • Ad-libs: pan them out wider (40–80%), give them more reverb and delay than the lead, and try band-pass "telephone" EQ (roughly 300 Hz–3 kHz) or distortion for contrast. That processing is a big part of the rage and hyperpop sound.
  • Harmonies: compress them harder than the lead so they stay steady, and tuck them under the hook as a pad rather than a second lead.

Compression Settings

Rap and melodic trap vocals need to be consistent. Every syllable should sit at a similar level over a loud beat. Two compressors in series usually sound more natural than one working hard.

StageStyleRatioAttackReleaseGain reduction
Compressor 1: peak controlFast, FET-style4:1 to 8:11–5 ms40–80 ms3–6 dB on peaks
Compressor 2: levelingSmooth, opto-style2:1 to 4:110–30 ms100–300 ms (or auto)2–4 dB steady
Parallel (optional)Heavy, on a send8:1 to 20:1FastFast10–15 dB, blended in low

A slower attack (10–30 ms) keeps consonants punchy. A faster attack smooths the delivery for melodic sections. Aggressive rage and hyperpop vocals often take more compression than you'd expect, and that's part of the style. Match the output gain so you're judging tone, not loudness.

De-Essing and Taming Sibilance

Compression and air boosts both bring out "s", "sh" and "t" sounds. Put a de-esser after the compressors and the high shelf:

  • Frequency: usually 5–8 kHz for male voices and 6–10 kHz for female voices. Use listen/solo mode to find the harshest band.
  • Reduction: 3–6 dB on the worst sibilants. More than that starts to make the artist sound lispy.
  • Split-band mode keeps the rest of the voice untouched. Wideband sounds more natural on softer sibilance.

For extreme cases, lower individual "s" sounds with clip gain before the chain. It's tedious, but nothing sounds cleaner.

Autotune and Pitch Correction Settings

Pitch correction has two uses: a transparent fix, or an audible effect that's part of the genre. Either way, set the key and scale to match the beat. If the key isn't listed, find the 808's root note or use a key-detection tool. The wrong scale is the most common reason an autotuned vocal sounds broken.

GoalRetune speedHumanize / flexNotes
Hard, robotic trap / rage effect0–10 ms (fastest)0Instant snapping between notes. Classic melodic trap sound.
Melodic but natural15–30 msLow to mediumCorrects clearly while keeping slides and vibrato
Transparent correction40–80 msMedium to highMostly invisible. Better yet, edit notes graphically.

Put pitch correction first in the chain, before EQ and compression, so it reads the cleanest signal. For hyperpop-style vocals, a formant shift of +1 to +3 semitones on doubles or a pitched-up duplicate an octave higher, blended low, gives that bright, synthetic character.

Heads-up: Hard-tune can't fix a performance that's far off key. It will snap to the wrong note. If a line glitches, retrack it or fix it graphically instead of fighting the plugin.

Reverb and Delay Sends Synced to BPM

Use reverb and delay on sends, not as inserts. That way several vocal tracks share one space and you control the wet level with a single fader. Delay times synced to the beat keep echoes in the groove instead of cluttering it. The formula for a quarter note is 60,000 ÷ BPM in milliseconds.

BPM1/4 note1/8 noteDotted 1/81/16 note
120500 ms250 ms375 ms125 ms
140428.6 ms214.3 ms321.4 ms107.1 ms
150400 ms200 ms300 ms100 ms
160375 ms187.5 ms281.3 ms93.8 ms

Delay send

  • 1/4 or dotted 1/8 synced to tempo, feedback 15–30%.
  • Filter the returns: high-pass at ~300 Hz, low-pass at ~6–8 kHz so repeats sit behind the dry voice.
  • Automate "delay throws" on the last word of a line instead of leaving the send up all the time.

Reverb send

  • A plate or room with a 0.8–1.8 s decay for verses. Longer (2–3 s) halls work for melodic hooks.
  • Pre-delay of 20–40 ms, or a 1/32–1/64 note, keeps consonants clear before the tail blooms.
  • High-pass the return at 200–400 Hz and low-pass around 8–10 kHz.
  • Put a de-esser or compressor before the reverb, sidechained to the dry vocal, so the space "breathes" between phrases.

Ducking the Beat Under the Vocal

This is the most important step with a 2-track. Because the beat already fills the whole spectrum, the vocal has to borrow space from it, especially in the 1–5 kHz range where both the voice's intelligibility and most lead synths live.

  1. Static EQ carve: cut 1–3 dB on the beat bus with a wide bell around 2–4 kHz. Subtle, but often enough.
  2. Dynamic EQ sidechain (best option): put a dynamic EQ on the beat bus, sidechain it from the lead vocal, and set a band around 1.5–4 kHz to dip 2–4 dB only while the vocal is present. Attack about 10 ms, release 100–200 ms.
  3. Multiband sidechain compression: same idea, using the mid band of a multiband compressor.
  4. With trackouts: sidechain only the melody or lead stem, and leave the drums and 808 alone so the groove stays punchy.

Done right, nobody hears the beat ducking. They just hear a vocal that sits inside the beat instead of on top of it.

Vocal Bus, Master and Loudness Targets

On the vocal bus, add gentle glue compression (2:1, 1–2 dB of reduction) and, if needed, light saturation for density. Then check the full mix in mono, on headphones, on earbuds and on a phone speaker.

Reference against released songs

Pick two or three released songs in the same style and import them into your session, turned down to match your mix's loudness so you compare tone, not level. Flip between them and your mix, paying attention to how loud the vocal sits relative to the kick and snare, how bright the top end is, and how much reverb you can actually hear. Most home mixes have vocals either too loud and dry or too quiet and washed out. Referencing catches both. Take breaks every 45–60 minutes, because your ears adapt to harshness and low-end buildup faster than you think.

On the master, a limiter brings the song to release level. Streaming platforms normalize playback, so pushing louder than necessary just costs dynamics. Common targets:

DestinationIntegrated loudnessTrue peak ceiling
Streaming reference (Spotify, YouTube normalization)around -14 LUFS-1.0 dBTP
Apple Music normalization referencearound -16 LUFS-1.0 dBTP
Competitive trap / rage master-9 to -7 LUFS-1.0 dBTP

Much of modern trap and rage is mastered louder than -14 LUFS for density and character. That's a stylistic choice, not a requirement. Whatever you choose, keep the true peak at -1 dBTP or below so the lossy encoding on streaming services doesn't clip. For a release-ready result, our mixing and mastering service handles the whole chain, and you can move on to our guide on releasing a song made on a type beat.

Next Steps: Beats, Stems and Pro Mixing

The quickest mix upgrade is better source material. Find your next instrumental in our beat catalog, and consider a Trackout Lease for control over individual parts. If you already own a lease and need separate stems, an arrangement edit or a key change, our beat customization and stems service can help. Want to learn to mix yourself? Book production lessons. Or look through all our services and send us a message through the contact form to get started.

Frequently asked questions

Can I mix vocals well on just a 2-track beat?

Yes. Turn the beat down to about -6 dBFS, process the vocal carefully, and use a dynamic EQ on the beat sidechained to the vocal around 1.5-4 kHz so the voice gets its own space. Trackouts give you more control.

What autotune retune speed should I use for trap?

For the hard, robotic trap and rage effect, use the fastest retune speed (roughly 0-10 ms) with humanize off. For a more natural melodic sound, try 15-30 ms. Always set the key and scale to match the beat.

How loud should my final master be?

Streaming services normalize to roughly -14 LUFS, but many trap and rage masters land at -9 to -7 LUFS by choice. Keep the true peak at -1 dBTP or lower either way.

TECHNOLOGY BEATS logo
TECHNOLOGY BEATS

Independent producer making hyperpop, rage and plugg type beats since 2020. About us · Browse beats

Keep reading

Related articles

All articles
—
Synth sketch of tempo & groove. Tagged previews on BeatStars.