Silence Remover
Automatically find and strip leading, trailing or internal silence — trim the dead air off the ends, shorten long pauses to a natural length, or remove every gap outright, with a live waveform showing exactly what's detected before you commit.
Quick Tips for Clean Results
Features
Three Ways to Handle a Silent Gap
"Remove the silence" turns out to mean different things depending on where that silence sits and why it's there, which is why this tool splits the job into three distinct modes rather than one blunt button. Trim Ends Only is the most conservative — it only touches silence at the very start and end of the file, the dead air before someone starts talking and after they finish, and leaves every pause in between exactly as recorded. It's the right choice when the recording's internal pacing is already fine and the only problem is a few seconds of nothing at either edge.
Remove All Silence sits at the opposite end — every detected gap, start, middle or end, gets deleted and what remains is spliced back together. That's genuinely useful for something like a raw multi-take recording where you only want the parts someone was actually speaking, back to back, with no dead air anywhere. It's a poor fit for anything meant to sound like natural conversation, though, since real speech relies on pauses for pacing and deleting all of them tends to sound rushed and mechanical.
Why Shortening Beats Deleting for Spoken Content
Shorten Long Pauses exists for the large middle ground between those two extremes, and it's the mode built to be the default choice for interviews, podcasts and voice memos. Instead of deleting a pause, it caps how long any single one is allowed to run — a six-second gap while someone gathers their thoughts gets cut down to whatever target length you set, commonly around 0.4 seconds, while genuinely short natural pauses between sentences are left untouched entirely, since they're already shorter than the target and don't need shortening at all.
The result keeps the actual rhythm of speech intact — sentences still have breathing room, changes of speaker still have a beat before the next one starts — while the specific stretches of true dead air, someone scrolling their notes or a long silent think, get compressed down to something that reads as a normal pause rather than an awkward hole. This is close to how professional podcast-editing software handles the same problem, and it's the mode most people reach for once they've compared it against a fully-deleted version and noticed how unnatural the rushed pacing sounds.
How the Detection Actually Works
Underneath all three modes is the same analysis step: the file gets scanned in tiny 10-millisecond windows, and each window's loudness is measured and compared against the Silence Threshold slider. Windows quieter than that threshold are marked silent, and a run of consecutive silent windows only becomes a removable gap once it's stayed silent for at least the Minimum Pause duration you've set — that second control exists specifically so a brief, completely normal dip in volume between two words doesn't get mistaken for a pause worth acting on.
Both sliders matter because "silent" isn't an absolute, fixed number — it depends entirely on how the recording was made. A quiet room and a good microphone can produce genuinely near-zero silence between phrases, so a strict threshold works well there. A phone recording made near traffic, an air conditioner, or a fan carries a real noise floor even during the "silent" parts, and a threshold set too strict will find nothing at all — which is exactly why the waveform's red shading updates live as the sliders move, so it's obvious at a glance whether the current settings are actually catching the gaps that are really there.
Who Actually Needs This
Podcasters trim the long pause at the very start of a raw recording before a guest realizes they're live, and shorten the thinking pauses scattered through the conversation without cutting them out entirely. Video creators clean up a voiceover take that has a few seconds of dead air between takes glued together in one file. Someone transcribing an interview strips the silence first so the audio moves faster through the parts that actually have speech in them. And plenty of people just have a voice memo with 20 seconds of pocket noise before anyone starts talking.
A podcaster with a two-hour raw recording full of thinking pauses runs it through Shorten Long Pauses first to tighten the pacing, then boosts what's left with the Volume Booster & Normalizer so every segment sits at a consistent, comfortable loudness. Someone assembling three separately-recorded voice memos trims the dead air off each one here, then brings the cleaned clips into the Audio Merger to join them into a single continuous file — a smoother result than merging first and having to hunt for silence buried in the middle of the combined recording afterward.
What Automatic Silence Removal Can't Fix
Detection works on loudness alone, not on meaning — it can't tell the difference between a genuinely dead pause and someone speaking unusually softly, so a very quiet speaker can end up with parts of actual speech flagged as silence at an aggressive threshold. Previewing the red-shaded waveform before applying anything is the fastest way to catch this: if shading appears over something that clearly isn't silence when you listen back, the threshold needs loosening, not tightening.
It also isn't a noise-reduction tool — a recording with constant background hum or hiss during its "quiet" parts still has that noise there after silence removal, because the noise itself was never truly below the threshold in the first place. If a recording needs more precise, manual control over exactly where a cut happens rather than automatic threshold-based detection, the Audio Trimmer & Cutter handles that instead, with draggable handles for picking an exact range by hand.