How to Mix a Podcast Episode

Mixing a podcast episode means balancing every track (voices, music, sound effects) so they sit together cleanly and play at a consistent, comfortable volume. The core moves are simple: clean the audio, level the voices, gently compress, add a touch of EQ, set music under speech, then match a loudness target before you export. You do not need a fancy room or expensive plugins to do this well, just a careful ear and a repeatable order of steps.
This article walks through that order using free and paid tools, with real settings you can copy. If you would rather skip the technical part entirely, our team at Timelapse Studio in Johar Town handles mixing as part of editing, but it helps every host to understand what is actually happening to their sound.
What mixing actually does
Recording captures the raw sound. Editing removes the parts you do not want (long pauses, mistakes, filler). Mixing is the stage where you make what is left sound good together. Think of it as the difference between three people talking at random volumes and three people who all sound like they are sitting at the same table with the same microphone.
A good mix does four things. It makes every speaker roughly the same loudness. It removes hiss, hum and room echo that distract the listener. It sets background music and stings so they support speech instead of fighting it. And it brings the whole episode to a loudness level that matches what platforms like Spotify, Apple Podcasts and YouTube expect.
If your raw recording is weak, mixing can only do so much. Getting clean tracks at the source matters more than any plugin, which is why studio quality audio starts with good microphone technique and a treated room.
Get your tracks ready first
Before you touch a single dial, organise your session. Mixing goes far faster when each speaker is on a separate track, which is the main reason to record everyone on separate tracks rather than one combined file.
- Import each voice onto its own track and label it (Host, Guest, Music, SFX).
- Listen through once at a low volume and note the obvious problems: a loud guest, a quiet host, a buzz, a phone vibration.
- Do your edits now. Cut mistakes, remove dead air, tighten gaps. Never mix and then edit, because every edit can change the balance you just set.
- Make a rough copy of the session so you can always go back if a setting goes wrong.
Once the edit is locked, you are ready to mix in the same order every time.
The mixing order, step by step
Always work in this sequence. Each step depends on the one before it, so jumping around creates more work.
| Step | What you do | Why it matters |
|---|---|---|
| 1. Clean | Remove noise, hum and clicks | A clean signal makes every later step easier |
| 2. Level | Match the volume of each voice | Listeners should not reach for the dial |
| 3. Compress | Even out loud and quiet moments | Keeps speech steady and present |
| 4. EQ | Shape the tone of each voice | Removes mud, adds clarity |
| 5. Music | Place and duck background audio | Music supports, never competes |
| 6. Loudness | Match a target and export | Plays at the right volume everywhere |
Step 1: Clean the audio
Start by removing anything that is not voice. Use a noise reduction tool to take out steady hiss or air conditioner hum, but go gently. Too much makes voices sound robotic and underwater. For one-off clicks, mouth noises and bumps, zoom in and cut or fade them manually. If you have constant echo from an untreated room, that is hard to fix after the fact, so it pays to reduce echo and background noise before you record.
Step 2: Level the voices
This is the single biggest improvement most episodes need. Your goal is for the host and guest to feel equally present. Use gain or volume automation to bring quiet sections up and pull loud bursts down. Aim for each voice to sit around -16 to -20 dB on average before compression. Do not chase a perfect number here, just get the speakers in the same ballpark so the compressor in the next step has an easy job.
Step 3: Compress each voice
Compression narrows the gap between the loudest and quietest parts of a track. For spoken voice, a gentle setting works best:
- Ratio: 2:1 to 3:1
- Attack: 10 to 20 ms (so the front of words stays natural)
- Release: 100 to 200 ms
- Threshold: set so the loudest words pull the meter down by 3 to 6 dB
You want the compressor working on peaks, not flattening everything. If the voice sounds squashed or pumping, ease off the ratio or raise the threshold.
Step 4: EQ for clarity
Equalisation shapes the tone. Three moves cover most voices. First, a high-pass filter around 80 to 100 Hz removes rumble and desk thumps you do not need. Second, a small cut around 200 to 400 Hz reduces boxy or muddy build-up. Third, a gentle lift around 3 to 6 kHz adds presence and makes words easier to understand. Make small changes, 2 to 3 dB at a time, and always compare against the original.
Step 5: Place the music
Intro and outro music can be loud and full. The moment a voice comes in, music must drop well underneath, usually 15 to 20 dB below the speech. Use volume automation or a ducking tool so the music fades down when someone talks and back up in the gaps. The test is simple: if you ever strain to hear a word over the music, the music is too loud.
Step 6: Match a loudness target
Platforms normalise loudness, so an episode that is too quiet gets turned up (bringing noise with it) and one that is too loud gets turned down. The widely used target for podcasts is around -16 LUFS for stereo and -19 LUFS for mono, with true peaks kept below -1 dB. Most editing apps have a built-in loudness meter, and free tools like Auphonic can match a target automatically. Getting this right is one of the most useful loudness standards for podcasts to learn once and reuse forever.
Tools you can actually use
You do not need a costly setup to mix well. The choice depends on your budget and how much control you want.
- Audacity (free): Handles noise reduction, EQ, compression and loudness for audio-only shows. Slower for big sessions but capable.
- GarageBand (free, Mac): Friendly for beginners, good built-in effects.
- Reaper (cheap licence): Powerful, lightweight, popular with serious podcasters.
- Adobe Audition or Logic Pro (paid): Full studio control with strong noise tools and automation.
- Auphonic (free tier): Online, automatic levelling and loudness matching. Great for fast turnaround.
If you produce a video podcast, you will still mix the audio with these same steps, then sync the finished audio back to your edited video. The picture changes nothing about the order of work.
Common mixing mistakes
A few habits ruin otherwise good episodes. Over-using noise reduction is the most common, leaving voices thin and metallic. Cranking the master volume instead of fixing levels per track is another, since it just makes the loud parts harsh. Adding heavy reverb or effects to a talking voice almost always sounds worse than leaving it dry. And mixing on cheap earbuds or laptop speakers hides problems you will only hear later on a phone or in a car, so check your mix on at least two different devices, including headphones.
The biggest mistake is mixing while exhausted. Ears tire fast. After 40 minutes everything starts to sound fine even when it is not. Take a break, then listen again the next day before you publish.
When to hand mixing to a studio
Mixing is a skill, and it takes hours per episode when you are learning. If your time is better spent recording and growing the show, it makes sense to hand it over. At Timelapse Studio we record at our space near Emporium Mall in Johar Town, Lahore, and offer editing and short-form clips as add-ons, so your audio is cleaned, levelled and loudness-matched without you opening a single plugin. We are open daily from 9 AM to 12 AM, and audio sessions start from Rs 5,000 per hour, with video from Rs 9,000 per hour and our Creator Pro option at Rs 15,000 per hour. Monthly and long-term bookings get up to 30 percent off, which suits anyone publishing on a regular schedule. You can compare options on our packages page.
For creators weighing the cost of doing it yourself versus booking help, our breakdown of podcast editing prices in Pakistan lays out what you should expect to pay.
Frequently asked questions
What is the difference between editing and mixing a podcast? Editing removes content you do not want, like mistakes, pauses and filler words. Mixing balances what remains so every voice and music track sounds clean, even and consistent in volume. Edit first, mix second.
What loudness should I export my podcast at? Aim for around -16 LUFS for stereo or -19 LUFS for mono, with true peaks below -1 dB. Most platforms normalise to roughly this range, so matching it keeps your show at a steady volume next to other podcasts.
Do I need expensive plugins to mix a podcast? No. Free tools like Audacity, GarageBand and Auphonic cover noise reduction, EQ, compression and loudness matching. A careful ear and a consistent order of steps matter far more than the price of your software.
Why do my voices sound different from each other? Usually it is a mix of different microphones, distances and room conditions, plus uneven levels. Recording each person on a separate track, then levelling, lightly compressing and EQing each voice individually, brings them into line.
How long does it take to mix one episode? When you are learning, expect one to two hours for a clean recording, longer if the audio has problems. Experienced editors and automated tools can do it in a fraction of that, which is why many hosts eventually hand it off.
Want clean, balanced episodes without the late nights in an editor? Book a session or message us and we will handle the mix for you.
Record at our Lahore studio
Audio and video podcast sessions with the gear and crew handled. Check the pricing or grab a slot.