跳转到主要内容
Workflows 5 min read

Guided Meditation Audio: Voice, Bed Music and Loudness Layers

Guided meditation looks like sleep audio’s sibling but builds on different logic: white noise wants “no events, no presence”, while meditation has a spine — the spoken guide. Voice stays clear and forward, the bed recedes into atmosphere, and the ending dissolves instead of stopping.

The essential difference from sleep audio

Sleep audio is flat: one even ambience looped all night, the fewer changes the better. Meditation is a line: settle in, guide the breath, scan the body, return — a voiced progression with the bed as its background color.

The craft follows: sleep audio’s core work is the loop point; meditation’s core work is loudness layering and fades.

Material: three components

The guided voice: spoken-word recorded close in a quiet room — voice quality is the life of meditation audio, and room reverb is the enemy. Bed music or ambience: low-level pure atmosphere (piano pad, distant rain), loop-friendly sources (selection criteria match sleep audio). Spaced silence: the breathing gaps between spoken passages — meditation’s rhythm lives half in silence.

Recording notes: 15–20 cm from the microphone, half your normal speaking pace, two-plus seconds between sentences — the “slow” of meditation is recorded in, not added later.

Loudness layering: voice forward, bed receded

Keep 10–15 dB between voice and bed: the voice clearly leads, the bed just perceptible. Practice: unify loudness separately (voice to normal spoken-word level, bed dropped 10–15 dB), then combine — the unification mechanics live in the volume guide.

The classic failure is bed at voice level: the guide drowns, the listener raises volume, the next voice entrance blasts. Unlayered meditation audio is simply unusable.

Fades: the etiquette of starting and ending

Open with 5–10 seconds of bed fade-in (an “entering” ramp), close with a 30–60 second slow fade-out — listeners at the end are in a relaxed state; a hard stop jolts them awake. After the fade, hold 2–3 seconds of true silence before the file ends.

Section transitions follow the same rule: the bed runs continuously under the voice (full-length bed, voice layered over), so the voice’s pauses never take the music down with them.

Structured assembly

The standard structure is a full-length bed with voice sections layered on: prepare the bed to target duration (loop or trim), then merge voice sections at their time points. The simpler alternative — inserting silence between voice sections and concatenating sequentially with the bed — avoids multi-track layering and is good enough.

Duration: 10–20 minutes per session (breath work 10, body scan 20) — too short never settles in, too long loses attention.

Formats and distribution

Publish as m4a or mp3 at 96–128 kbps — voice plus low-level bed has a simple spectrum; higher rates buy nothing. For sleep use, device looping takes over afterward: seam handling follows the sleep-audio approach (end fades covering the join).

Publishing to meditation apps or podcast platforms: re-check loudness against the platform’s stated ceiling — most podcast platforms set one.

Frequently asked questions