Tutorial · Research edition
Podcast Loudness and Export Settings: A Delivery Workflow
Prepare podcast audio with measured loudness, true-peak headroom, appropriate codec settings, metadata checks, and a decoded-file quality-control pass.
A podcast master can be clean in the editor and still fail after encoding. Loudness, true peak, channel count, sample rate, bit rate, metadata, and the host's own processing interact at the delivery boundary.
The responsible workflow keeps a high-quality master, creates a destination file from current requirements, measures before encoding, and listens to the decoded upload candidate—not only the timeline.
Separate mix decisions from delivery targets
Editing, EQ, dynamics, restoration, music balance, and fades create the master. Loudness normalization and peak control prepare that master for a distribution target. Pushing every phrase toward the target with heavy compression can reduce intelligibility and introduce distortion.
Measure the complete program, including intro, ads, music, and credits. A short loud segment does not describe integrated loudness; a safe sample peak does not guarantee safe true peak after encoding.
A documented target—not a universal law
Apple recommends preconditioning podcast audio to around -16 dB LKFS with ±1 dB tolerance and a true peak not exceeding -1 dBFS, using ITU-R BS.1770-5 calculations. Apple says this should occur before encoding because lossy compression may clip when true-peak headroom is insufficient.
Treat that as a current, traceable destination target. If a host, broadcaster, advertiser, or client supplies another specification, create and label the appropriate deliverable rather than forcing one file into every use.
Format and channel decisions
For RSS audio, Apple currently accepts MP3 or AAC and recommends AAC in MP4 for efficient streaming and accurate seeking. At 44.1 or 48 kHz, its recommended ranges are 64–128 kbps mono and 128–256 kbps stereo for both AAC and MP3.
Mono is efficient for a centered speech-only program. Stereo is required when music, spatial ambience, or deliberate left/right production carries meaning. Do not convert a mono voice into two channels and call it richer; choose based on content and compatibility.
Export ladder
Create files in a deliberate order.
- Archive master: lossless, full project sample rate/bit depth, no distribution codec.
- Delivery master: correct program edits, measured loudness and true peak, clean beginning/end.
- Encoded RSS file: host-supported MP3 or AAC, intentional channel count and bit rate.
- Platform video audio: aligned with the video workflow and platform requirements.
- Transcript, chapters, artwork, title, explicit flag, and episode metadata checked against the same version.
Decoded-file QC
Open the encoded file in a new session or player. Confirm duration, channel layout, start/end, ID3 or container metadata, chapter behavior when used, and absence of codec artifacts around sibilants, music, applause, and noise-reduced speech.
Measure again after encoding and compare against the master at matched playback level. Upload privately or as a draft when the platform permits, then inspect what listeners will actually receive.
Release record
Store the filename, checksum, duration, integrated loudness, loudness range when relevant, maximum true peak, sample rate, channel count, codec, bit rate, export preset version, and reviewer. This record makes a later correction reproducible rather than forensic.
Sources and verification
Claims were checked against the following first-party documentation on August 5, 2026. Product capabilities can change; verify current documentation before buying or changing a production system.
Frequently asked questions
What loudness should a podcast target?
Apple currently recommends about -16 dB LKFS with ±1 dB tolerance and true peak no higher than -1 dBFS. Treat this as a documented destination recommendation and follow any different host or client specification.
Should a spoken-word podcast be mono or stereo?
Mono is efficient when all meaningful content is centered speech. Use stereo when music, ambience, or intentional spatial production needs separate left and right channels.
Why check the encoded file?
Lossy encoding and metadata writing can introduce clipping, artifacts, channel mistakes, or version mismatches that are not audible in the lossless editing timeline.