Guided Meditation Audio: Voice, Bed Music and Loudness Layers
Guided meditation looks like sleep audio’s sibling but builds on different logic: white noise wants “no events, no presence”, while meditation has a spine — the spoken guide. Voice stays clear and forward, the bed recedes into atmosphere, and the ending dissolves instead of stopping.
The essential difference from sleep audio
Sleep audio is flat: one even ambience looped all night, the fewer changes the better. Meditation is a line: settle in, guide the breath, scan the body, return — a voiced progression with the bed as its background color.
The craft follows: sleep audio’s core work is the loop point; meditation’s core work is loudness layering and fades.
Material: three components
The guided voice: spoken-word recorded close in a quiet room — voice quality is the life of meditation audio, and room reverb is the enemy. Bed music or ambience: low-level pure atmosphere (piano pad, distant rain), loop-friendly sources (selection criteria match sleep audio). Spaced silence: the breathing gaps between spoken passages — meditation’s rhythm lives half in silence.
Recording notes: 15–20 cm from the microphone, half your normal speaking pace, two-plus seconds between sentences — the “slow” of meditation is recorded in, not added later.
Loudness layering: voice forward, bed receded
Keep 10–15 dB between voice and bed: the voice clearly leads, the bed just perceptible. Practice: unify loudness separately (voice to normal spoken-word level, bed dropped 10–15 dB), then combine — the unification mechanics live in the volume guide.
The classic failure is bed at voice level: the guide drowns, the listener raises volume, the next voice entrance blasts. Unlayered meditation audio is simply unusable.
Fades: the etiquette of starting and ending
Open with 5–10 seconds of bed fade-in (an “entering” ramp), close with a 30–60 second slow fade-out — listeners at the end are in a relaxed state; a hard stop jolts them awake. After the fade, hold 2–3 seconds of true silence before the file ends.
Section transitions follow the same rule: the bed runs continuously under the voice (full-length bed, voice layered over), so the voice’s pauses never take the music down with them.
Structured assembly
The standard structure is a full-length bed with voice sections layered on: prepare the bed to target duration (loop or trim), then merge voice sections at their time points. The simpler alternative — inserting silence between voice sections and concatenating sequentially with the bed — avoids multi-track layering and is good enough.
Duration: 10–20 minutes per session (breath work 10, body scan 20) — too short never settles in, too long loses attention.
Formats and distribution
Publish as m4a or mp3 at 96–128 kbps — voice plus low-level bed has a simple spectrum; higher rates buy nothing. For sleep use, device looping takes over afterward: seam handling follows the sleep-audio approach (end fades covering the join).
Publishing to meditation apps or podcast platforms: re-check loudness against the platform’s stated ceiling — most podcast platforms set one.