Why Your Podcast Needs Ducking

Listeners expect a polished audio experience where every word is clear and the background music adds energy without becoming a distraction. When you simply lower the music volume and leave it static, you risk either overpowering quiet speech or losing the musical momentum during pauses. Ducking — automatically reducing the music level when the host speaks — solves both problems. Sidechain compression is the industry-standard method to achieve this effect in any digital audio workstation (DAW).

By the end of this guide, you’ll know how to set up sidechain compression for ducking, which parameters matter most, and how to avoid common pitfalls. You’ll also discover advanced techniques for tailoring the effect to different podcast styles, from interview shows to narrative storytelling.

Understanding Sidechain Compression

Compression reduces the dynamic range of an audio signal by lowering the volume when it exceeds a threshold. Normally, a compressor works on the same track it’s processing. Sidechain compression adds a twist: the compressor listens to a different audio source (the sidechain input) to determine when to compress. For ducking, you place the compressor on your music track but feed your voice track into the sidechain input. Every time your voice exceeds the threshold, the compressor reduces the music volume. When you stop speaking, the music smoothly rises back to its original level.

This process mimics a live sound engineer manually pulling down the music fader during speech and pushing it up during pauses — but with precise, repeatable control. The result is a clean, professional mix where the voice always sits on top, yet the music never goes silent or sounds jarring.

Essential Compressor Parameters for Ducking

To get a natural ducking effect, you need to understand four key controls:

  • Threshold: The level at which the compressor starts acting. Set it so that the ducking engages only when your voice is present, not during very quiet breaths.
  • Ratio: Determines how much the music is reduced. Common ratios are 3:1 to 8:1. Higher ratios give more aggressive ducking but can sound unnatural if overdone.
  • Attack: How quickly the compressor reduces the music after the voice appears. A fast attack (1–5 ms) creates immediate ducking; slower attacks let a fraction of the music punch through before the ducking kicks in.
  • Release: How long the compressor takes to restore the music after the voice stops. A release of 100–300 ms is typical for a smooth rise; too short makes the music pop back abruptly, too long leaves the music suppressed during brief pauses.

Most compressors also offer a “sidechain filtering” option (high-pass or low-pass) that lets you duck only specific frequencies of the music — useful if you want the music’s bass to remain while only the mid-range ducks, for example.

Step-by-Step Setup in Your DAW

The exact menu names vary between DAWs, but the workflow is consistent. Below is a general guide followed by tips for popular software.

1. Prepare Your Tracks

Import your voice recording (or microphone input) onto one track, and your background music onto another. Name them clearly — “Voice” and “Music” — to avoid confusion.

2. Insert a Compressor on the Music Track

Load a compressor plugin that supports external sidechain input. Most modern DAW stock compressors support this (e.g., Ableton Live’s Compressor, Logic Pro’s Compressor, FL Studio’s Fruity Limiter, Pro Tools’ Dyn 3 Compressor/Limiter).

3. Enable Sidechain Input

Locate the sidechain section in the compressor interface. Typically you’ll find a dropdown menu where you can select the sidechain source. Choose your voice track. Some DAWs require you to route the audio manually via a bus; others automatically detect available tracks.

4. Set Initial Parameters

Start with these conservative settings and adjust by ear:

  • Ratio: 4:1
  • Attack: 5 ms
  • Release: 150 ms
  • Threshold: Set so the gain reduction meter shows 3–6 dB of reduction when you speak at normal loudness

5. Fine-Tune the Effect

Play a section where both voice and music play together. Listen for two things: is the music still distracting? And does the music return smoothly during natural pauses? Adjust the release time to match the rhythm of speech — longer for slower speech, shorter for fast-paced conversations. Lower the threshold if the ducking is too subtle; increase the ratio if you need more dramatic volume reduction.

6. Solo the Music (Optional)

To hear exactly how much and how fast the music ducks, temporarily solo the music track while the voice track is playing. This helps you evaluate whether the ducking is too aggressive or too slow. Reopen both tracks for final listening.

DAW-Specific Sidechain Tips

While the concept is universal, each DAW has slightly different steps. Here are pointers for the most common environments:

  • Ableton Live: On the compressor, click the “Sidechain” button. Under “Audio From,” select your voice track. Adjust the “EQ” section if you want frequency-specific ducking (e.g., only duck mids).
  • Logic Pro: Add a compressor to the music track. In the compressor’s sidechain menu (top right corner), choose the voice track. Enable “External” sidechain if needed. You can also use the “Side Chain” effect in Logic’s Channel EQ.
  • FL Studio: Insert Fruity Limiter on the music track. In the plugin window, set “Sidechain” to your voice track from the dropdown. Set the “Threshold” and “Ratio” in the limiter section. FL Studio also lets you route via mixer tracks — route both tracks to the same mixer insert and set the compressor’s sidechain input to the voice mixer track.
  • Pro Tools: Insert a Dyn 3 Compressor/Limiter on the music track. In the “Key Input” field, select the voice track as a bus or hardware input. Use the “Comp” section; expand “Key Filter” for frequency-specific ducking.
  • Audacity (free alternative): Audacity does not natively support real-time sidechain compression. You can achieve ducking by using a “Duck” effect from the “Effect” menu that manually lowers the music based on a voice track, but it renders the effect (not real-time). For real-time ducking in a free tool, consider Ocenaudio or install a VST sidechain compressor in Audacity via the “VST Enabler” extension.

Advanced Ducking Techniques

Once you’ve mastered the basic setup, you can refine the ducking to suit different podcast styles and creative goals.

Sidechain Frequency Filtering

By default, the compressor ducks the entire music signal. But you can instruct it to only duck specific frequencies — for example, only the mid-range (where voices live) while leaving bass and treble untouched. This preserves the music’s energy while still giving the voice clarity. Enable the “sidechain filter” on your compressor (often a high-pass or band-pass filter) and set it to around 300 Hz to 2 kHz — the typical voice range.

Multi-Band Ducking

For ultimate control, use a multi-band compressor on the music track and apply separate sidechain compression to different frequency bands. For instance, duck the 200–500 Hz range more aggressively to reduce boxiness, while only gently reducing the high end. This advanced technique requires careful setup but produces a very transparent mix.

Automated Ducking for Dynamic Sections

Some podcasts have varying energy levels — quiet introspective moments versus high-energy segments. Rather than using static compressor settings, you can automate the threshold or ratio of the compressor over time. For example, during a dramatic pause, you might lower the threshold so that even a whisper triggers lighter ducking, keeping the music softer. Many DAWs allow you to draw automation curves on compressor parameters.

Sidechain Compression on Other Elements

The same technique works for sound effects, jingles, or even ad breaks. You can duck any background element when the voice needs prominence. You can also duck the voice slightly against a loud music passage if you want the music to take over briefly — though in most podcasts, voice priority is standard.

Common Mistakes and How to Avoid Them

  • Over-compression: Setting the ratio too high (e.g., 20:1) or reducing the music level too much can make the ducking sound unnatural and pumpy. Stick to 3:1–8:1 and limit gain reduction to 6–10 dB.
  • Incorrect attack time: If the attack is too fast, the music disappears before the voice even starts, creating a hole. If too slow, the voice battles the music for the first syllable. A 3–8 ms attack usually works well.
  • Release too short: A release under 50 ms will make the music jump back up instantly, sounding jarring. Aim for at least 100 ms; longer releases (200–400 ms) are safer for natural speech rhythm.
  • Not monitoring in context: Always audition the mix with headphones and speakers. What sounds good on soloed music may sound wrong in the final mix.
  • Ignoring music dynamics: A highly dynamic music track (e.g., classical) may need a slower release to follow the natural ebb and flow. Heavily compressed pop music may need only gentle ducking to avoid sounding flat.
  • Forgetting headroom: Ensure your voice and music tracks have enough headroom (peaks around -6 dB) before applying compression, so the compressor doesn’t distort or cause artifacts.

Alternatives to Sidechain Ducking

Sidechain compression is the most popular method, but it’s not the only way to duck music:

  • Volume automation: Manually draw volume curves for the music track. This gives you total creative control but is time-consuming and may not react quickly to spontaneous speech.
  • Dynamic EQ: Use a dynamic equalizer that cuts only the frequencies occupied by the voice (e.g., 200–5 kHz) when the voice is present. This can sound more natural than full-band ducking because the music’s bass and high end remain unaffected.
  • Ducker plugins: Some plugins are purpose-built for ducking, like Waves Vocal Rider, Soundtheory’s “Uppercut” (though primarily a sidechain compressor), or the “Auto-Duck” feature in Adobe Audition. These often have simplified controls.
  • Multiband compression (already mentioned) and upward compression: Compressing the voice more can also help it cut through, but ducking music is usually more effective.

Practical Example: Ducking with a Voiceover

Imagine you have a podcast intro where a narrator speaks over an energetic synthwave track. Without ducking, you would have to lower the music volume globally, losing the drive. With sidechain compression set to a ratio of 5:1, attack 5 ms, release 200 ms, and threshold set to achieve 6 dB of reduction, the music ducks exactly when the narrator utters the first syllable and swells back during pauses. The result: the voice cuts through clearly, and the music feels alive.

For a calm interview segment with acoustic guitar, reduce the ratio to 3:1 and increase release to 300 ms for a more gentle, less obvious effect. Listen to the difference — the music becomes a soft bed rather than a competing element.

To deepen your understanding, explore these external resources:

Conclusion

Sidechain compression is a powerful, flexible tool for professional podcast audio. Once you set it up correctly, you’ll never want to go back to static music levels. The key is to experiment with attack, release, ratio, and threshold to find the sweet spot that complements your voice and music style. Whether you’re producing a daily news podcast or a storytelling series, ducked background music will keep your listeners engaged and focused on your content.

Start with the basic steps above, test with your own audio, and gradually incorporate advanced techniques like frequency filtering or automation. With practice, you’ll develop an instinct for the right settings — and your podcasts will sound polished, clear, and dynamic.