Podcasting and Audio Content: Plan, Record, Grow and MonetizeEditing and production · Lesson 11 of 18

Audio processing and loudness

Article · 7 min · 8 min lecture

Video lecture

Audio processing and loudness

13 chapters · about 8 min · full transcript

Coming soon

Chapter 1 of 13

Processing and loudness

  • Volume jumps between shows and voices
  • A six-step chain
  • LUFS and true peak
  • Music, licensing, a starting preset

The narrated lecture is in production

Every chapter is scripted and ready. Browse the chapters and read the full transcript now — the video will appear here when it’s published.

Chapters

From raw voice to polished sound

Processing makes voices clearer, more consistent and comfortable to listen to across earbuds, car speakers and phones. You don't need to be an audio engineer — a simple, consistent processing chain covers most shows.

A basic processing chain

Order matters. A common chain per voice track:

1. Clean-up      → Noise reduction (light), de-click, remove hum if present
2. EQ            → High-pass filter to remove rumble; gentle tone shaping
3. De-esser      → Tame harsh 's' and 'sh' sounds
4. Compression   → Even out volume differences
5. Leveling     → Balance voices against each other
6. Loudness      → Normalize the final mix to a target (LUFS) with a true-peak limit

Many editors offer presets or one-click "voice" processing that combine these steps. Use them as a starting point and adjust by ear.

Key concepts

  • Noise reduction: removes steady background noise (hiss, hum, AC). Overdoing it creates a "watery" or robotic sound — use lightly and fix noise at the source when possible.
  • EQ (equalization): adjusts frequencies. A high-pass filter (cutting low rumble below the voice) is almost always useful. Reduce "muddy" low-mids or harsh upper-mids gently; avoid extreme boosts.
  • Compression: reduces the gap between loud and quiet parts so listeners don't constantly adjust volume. Too much sounds squashed and tiring.
  • De-essing: reduces sharp sibilance.
  • Limiting: prevents peaks from exceeding a ceiling.

Loudness: LUFS and true peak

Loudness is measured in LUFS (Loudness Units relative to Full Scale), which reflects perceived loudness better than peak levels. Platforms normalize playback, so consistent loudness matters:

  • A widely used podcast target is around −16 LUFS integrated for stereo (roughly −19 LUFS for mono), with true peak no higher than about −1 dBTP. Apple's guidance for podcasts is around −16 LUFS with true peak below −1 dBFS.
  • Some platforms normalize to other levels (for example, some music-oriented platforms target around −14 LUFS). Aiming for the common podcast target and avoiding clipping works well across platforms.
  • Consistency across episodes and between voices matters as much as the exact number.

Most DAWs and podcast tools have loudness meters and normalization functions. Check current platform recommendations periodically.

Music and sound design

  • Intro/outro music: short, on-brand, mixed below voices when overlapping ("ducking").
  • Transitions/stingers: brief sounds to mark segments.
  • Licensing is essential: you need rights to any music you use. "I only used 10 seconds" is not a license. Options include royalty-free libraries with podcast/commercial licenses, commissioned music, or AI music tools whose terms permit commercial use (check carefully). Keep license records.
  • Avoid music under speech for long periods — it reduces clarity and can trigger content-ID claims on video platforms if not licensed.

Export settings

OutputCommon settings
Master archiveWAV, 24-bit, 44.1 or 48 kHz
Podcast distributionMP3 (e.g. 128 kbps mono for voice or higher for stereo/music-heavy), or as your host recommends
VideoAudio embedded per video platform recommendations (AAC common)

Add ID3 metadata (title, episode number, artwork) if your host doesn't handle it.

Listening checks

  • Listen on headphones, laptop speakers and a phone speaker.
  • Check the loudest moment (laughter) and quietest moment (soft-spoken guest).
  • Compare the new episode's loudness and tone with previous episodes.

Worked example: balancing a loud host and quiet guest

The host is close to a dynamic mic; the guest recorded on a laptop mic further away. Steps: light noise reduction on the guest track, high-pass filter on both, a gentle EQ cut in the guest's boxy frequencies, compression on both, leveling so both voices sit at similar perceived loudness, then final loudness normalization to −16 LUFS with a −1 dBTP limiter. Next time, the host sends the guest a USB mic and setup guide.

Hands-on: a starting preset you can adjust by ear

PER VOICE TRACK (starting points, adjust by ear)
High-pass filter      70-90 Hz (lower for deep voices)
Noise reduction       light; stop before voices sound "watery"
EQ                    small cut in "boxy" 200-500 Hz if needed; avoid big boosts
De-esser              reduce harsh 5-8 kHz only when "s" sounds sting
Compressor            ratio ~2:1 to 3:1, 3-6 dB gain reduction on louder phrases
MASTER
Loudness target       -16 LUFS integrated (stereo) / about -19 LUFS (mono)
True-peak limiter     ceiling -1 dBTP
Verify                measure the exported file, not only the project

Apple's published recommendation is overall loudness around -16 dB LKFS with a tolerance of plus or minus 1 dB, and true peak not above -1 dBFS, measured using ITU-R BS.1770. LUFS and LKFS describe the same measurement. Tools such as Auphonic or your editor's loudness function can normalize automatically, but always measure the final export.

Video podcasts and loudness

YouTube and Spotify normalize playback loudness, so a consistent, clean master is more important than being "loud". Use the same mastered audio for your audio and video versions so listeners get the same experience everywhere, and check that background music is licensed for video platforms as well as audio (content-ID systems can flag unlicensed tracks).

Common mistakes

  • Heavy noise reduction creating artifacts.
  • Extreme EQ boosts.
  • Over-compression that sounds squashed.
  • Inconsistent loudness between episodes.
  • Unlicensed music.

Summary

Use a simple processing chain — clean-up, EQ, de-ess, compression, leveling, loudness — applied gently; target consistent loudness (commonly around −16 LUFS stereo with true peak below −1 dBTP); license all music; export masters and distribution files properly; and check on multiple devices.

Key takeaways

  • A simple chain — clean-up, EQ, de-ess, compression, leveling, loudness — covers most shows.
  • Apply processing gently; fix noise at the source where possible.
  • A common podcast target is around −16 LUFS stereo (−19 mono) with true peak below about −1 dBTP.
  • License all music and keep records; 'only a few seconds' is not a license.

Check your understanding

Quick questions to lock in the lesson. They don’t count towards your certificate.

  1. What does LUFS measure?
  2. Which loudness target is widely used for stereo podcasts?
  3. You want to use 10 seconds of a popular song as intro music. What do you need?

Put it into practice

Build a processing preset in your editor using the six-step chain, process one episode, measure integrated loudness and true peak, and compare it on three different playback devices.

Enrol for free to save your progress

Reading is always free. Enrol to keep your place, take the final assessment and earn a verifiable certificate.