---
title: "Audio processing and loudness | Optimize All Academy"
description: "From raw voice to polished sound Processing makes voices clearer, more consistent and comfortable to listen to across earbuds, car speakers and phones…"
url: https://optimizeall.com/learn/podcasting-and-audio-content/audio-processing-and-loudness
updated: 2026-10-05
---

Podcasting and Audio Content: Plan, Record, Grow and Monetize · Editing and production · lesson 11 of 18 · 7 min

# Audio processing and loudness

## From raw voice to polished sound

Processing makes voices clearer, more consistent and comfortable to listen to across earbuds, car speakers and phones. You don't need to be an audio engineer — a simple, consistent processing chain covers most shows.

## A basic processing chain

Order matters. A common chain per voice track:

```
1. Clean-up      → Noise reduction (light), de-click, remove hum if present
2. EQ            → High-pass filter to remove rumble; gentle tone shaping
3. De-esser      → Tame harsh 's' and 'sh' sounds
4. Compression   → Even out volume differences
5. Leveling     → Balance voices against each other
6. Loudness      → Normalize the final mix to a target (LUFS) with a true-peak limit
```

Many editors offer presets or one-click "voice" processing that combine these steps. Use them as a starting point and adjust by ear.

## Key concepts

- **Noise reduction:** removes steady background noise (hiss, hum, AC). Overdoing it creates a "watery" or robotic sound — use lightly and fix noise at the source when possible.
- **EQ (equalization):** adjusts frequencies. A high-pass filter (cutting low rumble below the voice) is almost always useful. Reduce "muddy" low-mids or harsh upper-mids gently; avoid extreme boosts.
- **Compression:** reduces the gap between loud and quiet parts so listeners don't constantly adjust volume. Too much sounds squashed and tiring.
- **De-essing:** reduces sharp sibilance.
- **Limiting:** prevents peaks from exceeding a ceiling.

## Loudness: LUFS and true peak

Loudness is measured in **LUFS** (Loudness Units relative to Full Scale), which reflects perceived loudness better than peak levels. Platforms normalize playback, so consistent loudness matters:

- A widely used podcast target is around **−16 LUFS integrated for stereo** (roughly −19 LUFS for mono), with **true peak no higher than about −1 dBTP**. Apple's guidance for podcasts is around −16 LUFS with true peak below −1 dBFS.
- Some platforms normalize to other levels (for example, some music-oriented platforms target around −14 LUFS). Aiming for the common podcast target and avoiding clipping works well across platforms.
- Consistency across episodes and between voices matters as much as the exact number.

Most DAWs and podcast tools have loudness meters and normalization functions. Check current platform recommendations periodically.

## Music and sound design

- **Intro/outro music:** short, on-brand, mixed below voices when overlapping ("ducking").
- **Transitions/stingers:** brief sounds to mark segments.
- **Licensing is essential:** you need rights to any music you use. "I only used 10 seconds" is not a license. Options include royalty-free libraries with podcast/commercial licenses, commissioned music, or AI music tools whose terms permit commercial use (check carefully). Keep license records.
- Avoid music under speech for long periods — it reduces clarity and can trigger content-ID claims on video platforms if not licensed.

## Export settings

| Output | Common settings |
|---|---|
| Master archive | WAV, 24-bit, 44.1 or 48 kHz |
| Podcast distribution | MP3 (e.g. 128 kbps mono for voice or higher for stereo/music-heavy), or as your host recommends |
| Video | Audio embedded per video platform recommendations (AAC common) |

Add **ID3 metadata** (title, episode number, artwork) if your host doesn't handle it.

## Listening checks

- Listen on headphones, laptop speakers and a phone speaker.
- Check the loudest moment (laughter) and quietest moment (soft-spoken guest).
- Compare the new episode's loudness and tone with previous episodes.

## Worked example: balancing a loud host and quiet guest

The host is close to a dynamic mic; the guest recorded on a laptop mic further away. Steps: light noise reduction on the guest track, high-pass filter on both, a gentle EQ cut in the guest's boxy frequencies, compression on both, leveling so both voices sit at similar perceived loudness, then final loudness normalization to −16 LUFS with a −1 dBTP limiter. Next time, the host sends the guest a USB mic and setup guide.

## Hands-on: a starting preset you can adjust by ear

```text
PER VOICE TRACK (starting points, adjust by ear)
High-pass filter      70-90 Hz (lower for deep voices)
Noise reduction       light; stop before voices sound "watery"
EQ                    small cut in "boxy" 200-500 Hz if needed; avoid big boosts
De-esser              reduce harsh 5-8 kHz only when "s" sounds sting
Compressor            ratio ~2:1 to 3:1, 3-6 dB gain reduction on louder phrases
MASTER
Loudness target       -16 LUFS integrated (stereo) / about -19 LUFS (mono)
True-peak limiter     ceiling -1 dBTP
Verify                measure the exported file, not only the project
```

Apple's published recommendation is overall loudness around -16 dB LKFS with a tolerance of plus or minus 1 dB, and true peak not above -1 dBFS, measured using ITU-R BS.1770. LUFS and LKFS describe the same measurement. Tools such as Auphonic or your editor's loudness function can normalize automatically, but always measure the final export.

## Video podcasts and loudness

YouTube and Spotify normalize playback loudness, so a consistent, clean master is more important than being "loud". Use the same mastered audio for your audio and video versions so listeners get the same experience everywhere, and check that background music is licensed for video platforms as well as audio (content-ID systems can flag unlicensed tracks).

## Common mistakes

- Heavy noise reduction creating artifacts.
- Extreme EQ boosts.
- Over-compression that sounds squashed.
- Inconsistent loudness between episodes.
- Unlicensed music.

## Summary

Use a simple processing chain — clean-up, EQ, de-ess, compression, leveling, loudness — applied gently; target consistent loudness (commonly around −16 LUFS stereo with true peak below −1 dBTP); license all music; export masters and distribution files properly; and check on multiple devices.

## Video lecture: Audio processing and loudness

Lecture coming soon · 13 chapters · about 8 minutes. Read the full transcript below.

1. Processing and loudness
2. Why process?
3. Processing is photo prep
4. The six-step chain
5. Loudness targets
6. Example 1: loud host, quiet guest
7. Example 2: three shows, one standard
8. Watch me: a starting preset
9. Music and licensing
10. Exports and video
11. Mistakes + recap
12. Listening checks
13. Try this now

## Lecture transcript

### Processing and loudness

Have you ever listened to a podcast in the car, then switched to another show and had to grab the volume knob because it was suddenly twice as loud? Or struggled to hear a quiet guest on a train, then been blasted when the host laughed? Those are processing and loudness problems. And they're fixable with a simple, consistent chain. In this lesson, you'll learn a six-step processing chain, what LUFS and true peak actually mean, Apple's published loudness recommendation, how to handle music and licensing, and a starting preset you can adjust by ear.

### Why process?

Why does processing matter? Because people listen on earbuds, car speakers, laptop speakers and phones, often in noisy places. Processing makes voices clearer, more consistent and more comfortable across all of them. You don't need to be an audio engineer. A simple chain, applied gently and consistently, covers most shows. And consistency matters as much as any exact number. Listeners notice when one episode is much louder than the last, or when the guest is much quieter than the host.

### Processing is photo prep

Here's an analogy. Think of processing like preparing a photo for print. First you remove dust spots. That's noise reduction. Then you adjust the color balance. That's EQ. You soften any harsh highlights. That's de-essing. You bring the shadows and highlights closer together so details show. That's compression. You balance two photos on the same page. That's leveling. And finally you set the overall brightness to the printer's standard. That's loudness normalization. Do each step lightly, and the photo looks natural. Overdo any of them, and it looks fake.

### The six-step chain

Next, the chain, per voice track, in order. One, cleanup: light noise reduction, de-click and hum removal if needed. Two, EQ: a high-pass filter to remove low rumble, and gentle tone shaping. Three, de-essing, to tame harsh s and sh sounds. Four, compression, to even out loud and quiet phrases. Five, leveling, to balance the voices against each other. And six, loudness: normalize the final mix to a target, with a true-peak limit. Many editors offer one-click voice presets that combine these steps. Use them as a starting point, and adjust by ear.

### Loudness targets

Now loudness. LUFS, loudness units relative to full scale, measures perceived loudness better than peak levels. Platforms normalize playback, so consistent loudness matters. A widely used podcast target is around minus sixteen LUFS integrated for stereo, or about minus nineteen for mono, with true peak no higher than about minus one. Apple's published recommendation is overall loudness around minus sixteen LKFS, with a tolerance of plus or minus one, and true peak not above minus one dBFS, measured using an international standard called ITU-R BS seventeen seventy. LUFS and LKFS describe the same measurement. Most importantly, measure your exported file, not just the project.

### Example 1: loud host, quiet guest

A simple example. A host sits close to a dynamic mic. The guest recorded on a laptop mic, further away. The fix: light noise reduction on the guest track. A high-pass filter on both. A gentle EQ cut in the guest's boxy frequencies. Compression on both. Leveling so both voices sit at similar perceived loudness. Then final normalization to minus sixteen LUFS with a minus one limiter. And the most important fix of all: next time, the host sends the guest a USB mic and the setup guide.

### Example 2: three shows, one standard

Now a realistic scenario. A network of three shows in the UAE, run by one small team, gets listener complaints that episodes sound different every week. Each editor uses their own settings. So they standardize. One processing preset per show, saved as a template. One mastering target: minus sixteen LUFS integrated and minus one true peak, applied with Auphonic as the final step. And a rule: every export is measured, and the number goes into the publishing checklist. Within a month, complaints stop, and switching between their shows no longer means reaching for the volume.

### Watch me: a starting preset

Watch me build the starting preset from your lesson. Per voice track: high-pass filter at seventy to ninety hertz, lower for deep voices. Light noise reduction, stopping before voices sound watery. A small EQ cut in the boxy two hundred to five hundred hertz range, only if needed. A de-esser on the harsh five to eight kilohertz range, only when s sounds sting. A compressor at around two to one or three to one, with three to six decibels of gain reduction on louder phrases. Then on the master: a loudness target of minus sixteen LUFS integrated, and a true-peak limiter with a ceiling of minus one. I export, measure the exported file, and it reads minus sixteen point two. Within tolerance. Done.

### Music and licensing

Now music and sound design. Short, on-brand intro and outro music, mixed below voices when they overlap, which is called ducking. Brief stingers to mark segments. And licensing is essential. You need rights to any music you use. "I only used ten seconds" is not a license. Options include royalty-free libraries with podcast and commercial licenses, commissioned music, or AI music tools whose terms permit commercial use, checked carefully. Keep license records. Avoid music under speech for long stretches, because it hurts clarity. And check that your music is licensed for video platforms too, because content-ID systems can flag unlicensed tracks.

### Exports and video

Let's cover exports and video. Master archive: WAV, twenty-four bit, forty-four point one or forty-eight kilohertz. Podcast distribution: MP3, for example one hundred and twenty-eight kilobits per second mono for voice, or higher for stereo and music-heavy shows, or whatever your host recommends. Video: audio embedded per the platform's recommendations. Add metadata if your host doesn't. And use the same mastered audio for your audio and video versions, so listeners get the same experience everywhere. YouTube and Spotify normalize playback, so a clean, consistent master matters more than being loud.

### Mistakes + recap

Here are the common mistakes. Heavy noise reduction that creates watery artifacts. Extreme EQ boosts. Over-compression that sounds squashed and tiring. Inconsistent loudness between episodes. Measuring the project instead of the export. And unlicensed music. Let's recap. Use the six-step chain, applied gently. Fix noise at the source when possible. Target around minus sixteen LUFS for stereo, with true peak at or below minus one. License all music and keep records. Export masters and distribution files properly. And check on headphones, laptop speakers and a phone.

### Listening checks

One more practical habit: listening checks. Always listen on at least three devices: headphones, laptop speakers and a phone speaker. Check the loudest moment, usually laughter, and the quietest, usually a soft-spoken guest. Then compare the new episode's loudness and tone with your previous episode. If they sound like two different shows, something in the chain changed. These checks take ten minutes and catch most problems before listeners do.

### Try this now

Here's your try this now. Build a processing preset in your editor using the six-step chain and the starting settings in the lesson. Process one episode. Measure the integrated loudness and true peak of the exported file. Then listen on three devices. Save the preset as a template for every future episode. In the next lesson, we'll package your episode with titles, show notes, chapters and transcripts.

## Key takeaways

- A simple chain — clean-up, EQ, de-ess, compression, leveling, loudness — covers most shows.
- Apply processing gently; fix noise at the source where possible.
- A common podcast target is around −16 LUFS stereo (−19 mono) with true peak below about −1 dBTP.
- License all music and keep records; 'only a few seconds' is not a license.

## Try it

Build a processing preset in your editor using the six-step chain, process one episode, measure integrated loudness and true peak, and compare it on three different playback devices.

- [Previous: An efficient, ethical editing workflow](https://optimizeall.com/learn/podcasting-and-audio-content/editing-workflow)
- [Next: Show notes, chapters, transcripts and packaging](https://optimizeall.com/learn/podcasting-and-audio-content/show-notes-transcripts-and-packaging)
- [All lessons of Podcasting and Audio Content: Plan, Record, Grow and Monetize](https://optimizeall.com/learn/podcasting-and-audio-content)
