IT
OmnvertImage • Document • Network
Jul 18, 2026intermediate9 minaudio · silence · voice · editing · podcastSilence TrimmerMore guides for this tool

Remove Silence and Dead Air from Voice Recordings

Tune the silence threshold and minimum duration so you cut dead air without chopping breaths, word tails, or natural pauses.

Prerequisites

Supplies
  • A voice recording: lecture, memo, interview, or raw podcast track
  • Ten minutes to test a couple of settings before committing
Tools
  • Silence Trimmer

Step-by-step

  1. Load the recording and let it scan

    Upload your recording to the Silence Trimmer and let it analyze the levels. The tool works by comparing the loudness of every moment against a threshold you set, and marking anything quieter than that, for long enough, as silence to remove. Before you change any numbers, look at how much of the waveform is flat. A lecture recorded in a quiet room might be 20 percent dead air; a nervous first-take voice memo can be closer to 40 percent. That flat portion is your opportunity.

  2. Set the silence threshold in decibels

    The threshold is a dB level below which audio counts as silence. A common starting point is around -40 dB. Set it too high, say -25 dB, and the tool will treat quiet speech, soft word endings, and breaths as silence and delete them, leaving abrupt jumps. Set it too low, say -55 dB, and it only catches truly digital-silent gaps, missing the low room hum you wanted gone. The right value sits just above your room's background noise floor and just below the quietest speech you want to keep.

  3. Set a minimum silence duration

    This is the setting people forget, and it's what separates a natural result from a robotic one. Minimum duration tells the tool to only remove a quiet stretch if it lasts longer than, say, 500 milliseconds or a full second. Without it, the tool snips every tiny gap between words, including the micro-pauses that make speech sound human. With a 700 millisecond minimum, normal pauses survive untouched and only the genuine dead air, the three-second gap where you were thinking, gets collapsed.

  4. Decide how much silence to leave behind

    Removing a gap entirely often sounds worse than shortening it. Many tools let you keep a padding of, say, 200 to 300 milliseconds around each retained section so sentences don't slam into each other. Think about rhythm. Human speech breathes; a talk with every pause surgically removed feels breathless and stressful to listen to, even if you can't articulate why. Aim to shorten long gaps to a natural beat rather than erase them to zero. The goal is tighter, not airless.

  5. Treat edges differently from internal gaps

    Leading and trailing silence, the dead time before you start talking and after you stop, can be cut hard with no downside. Nobody misses the four seconds of you reaching for the record button. Internal gaps are the delicate ones, because that's where over-aggressive settings do damage. A good workflow is to trim the head and tail aggressively, then use gentler threshold and duration settings for the middle where the actual conversation lives. Two passes with different settings beats one blunt pass across the whole file.

  6. Preview the joins, not just the whole file

    After processing, don't just play the file start to finish. Jump specifically to the points where cuts were made and listen to each join. A bad join sounds like a word stepping on the next one, or a breath cut in half so it becomes a little gasp. If you hear those, your threshold is too aggressive or your padding is too small. Adjust and re-run. Checking the seams directly saves you from publishing a lecture that sounds subtly chopped throughout.

  7. Do this before loudness normalization

    Order matters. Remove silence first, then run loudness normalization. Normalization measures the average loudness of the whole file to set its target level. If you normalize first, all those silent gaps drag the average down and the tool over-boosts to compensate, which can push your speech too hot. Trim the dead air first so the loudness measurement reflects only the parts people actually hear. See the loudness tutorial for the exact targets.

What over-aggressive trimming destroys

  • Breaths: a natural inhale before a sentence gives listeners a cue that a new thought is coming. Cut them all and speech feels frantic.
  • Word tails: soft endings on words like 'time' or 'go' fade below the threshold. Cut them and the word sounds bitten off.
  • Thinking pauses: a one-second beat before an answer reads as thoughtfulness. Remove every one and the speaker sounds like a machine gun.

Good starting settings by recording type

For a clean lecture in a quiet room, try a threshold near -42 dB with a 600 to 800 millisecond minimum and 250 milliseconds of padding. For a phone voice memo with more background noise, raise the threshold to around -35 dB so the hum counts as silence, but be more generous with minimum duration to protect quiet speech. There's no universal number; the correct settings depend on your microphone, your room, and how close you sat. Test on a two-minute slice before committing to a two-hour file.

Keep an untouched original

Silence removal is destructive: once a gap is gone, you can't put a natural pause back by hand without it sounding fake. Always keep the raw recording so you can re-run with softer settings if the first pass cut too deep.

Related