Learn · Vocal mixing

How to get warm tube vocals for free

"Warm" is the most requested and least defined word in vocal mixing. Here's what it actually means, and a five-step chain that gets you there without spending anything.

Kalu Audio · October 9, 2026 · 7 min read

What "warm" really means

When engineers call a vocal warm, they usually mean three things at once:

  • Smooth highs. No harsh "S" sounds or brittle top end that makes the listener reach for the volume knob.
  • A full low-mid. The body of the voice (roughly 150 to 400 Hz) is present without getting muddy.
  • Gentle harmonic density. Low-order harmonics, especially the 2nd, thicken the tone and help the voice sit forward in the mix, even at the same level.

That last point is what tube gear is famous for. A single-ended triode stage, the classic tube preamp circuit, distorts asymmetrically, which produces mostly 2nd harmonic: an octave above the note, musically consonant, heard as richness rather than distortion. Every step below serves one of those three goals.

The chain, in order

Order matters more than any single setting. This is the sequence most vocal engineers converge on, and the one we built into Triodia:

  1. De-esser: tame sibilance first.
  2. Tube compressor: control dynamics and add harmonics.
  3. Echo: a sense of space that keeps the vocal up front.
  4. Plate reverb: depth and polish.
  5. Output: level-match before you judge.

Before you start, check your gain staging. Analog-modeled plugins react to input level just like hardware. Feed them a vocal averaging around −18 dBFS with peaks well below 0. Too hot and everything saturates; too quiet and nothing does.

Step 1: De-ess before you compress

Compression lowers the loud parts of a vocal, which means everything else, including sibilance, comes up relative to the body of the voice. De-ess first and you hand the compressor a smoother signal.

  • Most sibilance lives between 5 and 9 kHz. Start at about 6.5 kHz and sweep until the "S" sounds react most.
  • Use only as much reduction as needed to stop the S's from jumping out. Overdo it and the singer starts to lisp.
  • Listen in context, not soloed. A slightly bright vocal often sounds right in a full mix.

Step 2: Tube compression (this is where the warmth comes from)

A tube compressor does two jobs: it evens out the performance and it runs the voice through a saturating stage. Good starting points:

  • Ratio 3:1 to 4:1 for most pop and rock vocals.
  • Attack 10 to 30 ms. Faster grabs consonants and flattens the delivery; slower lets transients through for more punch.
  • Release 100 to 200 ms, or an adaptive, opto-style release that follows the phrasing.
  • Lower the threshold until you see 3 to 6 dB of gain reduction on the loudest phrases.
  • Drive sets how hard the tube stage is pushed. Raise it until the voice thickens, then back off a touch. If you can clearly hear "distortion", you've gone too far for a lead vocal.

The tube type changes the character a lot. A high-gain 12AX7 saturates early and sounds obviously warm; a low-gain 12AU7 stays clean with subtle color. We compare nine tubes, with measured numbers, in 12AX7 vs 12AU7 on vocals.

Step 3: Echo that stays out of the way

A delay gives a vocal size without pushing it back like a big reverb does. The trick is to make the repeats darker than the dry voice:

  • Time the delay to the song: a quarter or eighth note, or a short slapback of 80 to 150 ms for a vintage feel.
  • Roll off the highs of the repeats (around 4 to 6 kHz) so they never compete with the consonants.
  • Keep feedback low (two or three audible repeats) and the mix between 10 and 25%.
  • Ping-pong mode spreads the repeats across the stereo field and leaves the center for the lead.

Step 4: Plate reverb for depth

Plates have been the go-to vocal reverb since the 1950s: dense, bright and smooth, without the obvious room reflections of a hall. Three settings do most of the work:

  • Pre-delay 20 to 80 ms. This gap keeps the dry voice intelligible before the reverb blooms.
  • Low cut at 200 to 400 Hz. Reverb on the low end is the fastest route to mud.
  • Mix low. Raise it until you hear it, then pull it back a little. You should miss it when it's bypassed, not notice it when it's on.

Step 5: Level-match before you judge

Louder always sounds better, so a processed vocal that is 2 dB hotter will win every comparison. Adjust the output so the processed and bypassed versions are equally loud, then decide whether the chain is actually helping.

Doing it all with one free plugin

Triodia is our free VST3 for Windows that runs this exact chain in one window. A quick way in:

  1. Load it on the vocal track and pick the preset closest to your style: Pop, Rock, Ballad, Soul, Rap, Podcast, Voice-Over and more. All presets are level-matched, so you're comparing tone, not volume.
  2. Lower THRESH until the VU, in gain-reduction mode, moves a few divisions on the loud phrases.
  3. Choose a tube: ROCK (12AX7) for obvious warmth, SMOOTH (12AU7) for subtle color, GLUE or SILK for soft, vari-mu style compression. Raise DRIVE to hear the difference.
  4. Set the echo and plate MIX to taste, and watch the tube on the panel glow as the voice pushes it.
Free download

Try the whole chain in Triodia

De-esser, tube compressor with 9 modeled tubes, echo and plate reverb. VST3 + Standalone for Windows.