ToollyX
auto-detect · runs locally · free

Silence Remover

Automatically find and strip leading, trailing or internal silence — trim the dead air off the ends, shorten long pauses to a natural length, or remove every gap outright, with a live waveform showing exactly what's detected before you commit.

ModesTrim/Shorten/Remove
DetectionThreshold + duration
FormatsMP3/WAV/OGG
CostFree
728 × 90 — Leaderboard Ad
Drop an audio file here, or click to browse
MP3, WAV, OGG, M4A, AAC, FLAC
728 × 90 — Leaderboard Ad

Quick Tips for Clean Results

1Watch the red-shaded waveform, not just the numbers. It shows exactly which stretches count as silence at your current settings, so you can confirm it's catching the right gaps before applying anything.
2Start with Shorten Long Pauses for spoken content. It removes dead air without making speech sound unnaturally rushed — reach for Remove All Silence only when you specifically want every gap gone.
3Loosen the threshold for noisy recordings. A phone recording with background hum rarely has true digital silence, so a strict threshold like -55dB may detect nothing at all — try -30dB or -25dB instead.
4Keep Minimum Pause above 0.2s. Shorter than that starts catching the natural micro-gaps between individual words, which shouldn't be treated as removable dead air.

Features

Automatic detection
Finds silent stretches by threshold and minimum duration — no manual scrubbing through the waveform.
3 handling modes
Trim ends only, shorten long pauses, or remove every gap — pick whichever matches the actual goal.
Live waveform highlighting
Detected silence is shaded directly on the waveform as you adjust the threshold, before you apply anything.
Adjustable sensitivity
Threshold and minimum pause duration both tune to the specific recording, not a one-size-fits-all default.
Click-free crossfaded cuts
Every splice blends smoothly so removing or shortening silence never introduces an audible pop.
Original vs. current A/B
Switch instantly between the source file and the edited version to compare before exporting.
Iterative editing
Apply, listen, adjust settings, and apply again — with a one-click reset back to the original.
Export to MP3, WAV or OGG
Pick the format the cleaned-up file actually needs, with adjustable MP3 bitrate.

Three Ways to Handle a Silent Gap

"Remove the silence" turns out to mean different things depending on where that silence sits and why it's there, which is why this tool splits the job into three distinct modes rather than one blunt button. Trim Ends Only is the most conservative — it only touches silence at the very start and end of the file, the dead air before someone starts talking and after they finish, and leaves every pause in between exactly as recorded. It's the right choice when the recording's internal pacing is already fine and the only problem is a few seconds of nothing at either edge.

Remove All Silence sits at the opposite end — every detected gap, start, middle or end, gets deleted and what remains is spliced back together. That's genuinely useful for something like a raw multi-take recording where you only want the parts someone was actually speaking, back to back, with no dead air anywhere. It's a poor fit for anything meant to sound like natural conversation, though, since real speech relies on pauses for pacing and deleting all of them tends to sound rushed and mechanical.

Why Shortening Beats Deleting for Spoken Content

Shorten Long Pauses exists for the large middle ground between those two extremes, and it's the mode built to be the default choice for interviews, podcasts and voice memos. Instead of deleting a pause, it caps how long any single one is allowed to run — a six-second gap while someone gathers their thoughts gets cut down to whatever target length you set, commonly around 0.4 seconds, while genuinely short natural pauses between sentences are left untouched entirely, since they're already shorter than the target and don't need shortening at all.

The result keeps the actual rhythm of speech intact — sentences still have breathing room, changes of speaker still have a beat before the next one starts — while the specific stretches of true dead air, someone scrolling their notes or a long silent think, get compressed down to something that reads as a normal pause rather than an awkward hole. This is close to how professional podcast-editing software handles the same problem, and it's the mode most people reach for once they've compared it against a fully-deleted version and noticed how unnatural the rushed pacing sounds.

How the Detection Actually Works

Underneath all three modes is the same analysis step: the file gets scanned in tiny 10-millisecond windows, and each window's loudness is measured and compared against the Silence Threshold slider. Windows quieter than that threshold are marked silent, and a run of consecutive silent windows only becomes a removable gap once it's stayed silent for at least the Minimum Pause duration you've set — that second control exists specifically so a brief, completely normal dip in volume between two words doesn't get mistaken for a pause worth acting on.

Both sliders matter because "silent" isn't an absolute, fixed number — it depends entirely on how the recording was made. A quiet room and a good microphone can produce genuinely near-zero silence between phrases, so a strict threshold works well there. A phone recording made near traffic, an air conditioner, or a fan carries a real noise floor even during the "silent" parts, and a threshold set too strict will find nothing at all — which is exactly why the waveform's red shading updates live as the sliders move, so it's obvious at a glance whether the current settings are actually catching the gaps that are really there.

728 × 90 — Leaderboard Ad

Who Actually Needs This

Podcasters trim the long pause at the very start of a raw recording before a guest realizes they're live, and shorten the thinking pauses scattered through the conversation without cutting them out entirely. Video creators clean up a voiceover take that has a few seconds of dead air between takes glued together in one file. Someone transcribing an interview strips the silence first so the audio moves faster through the parts that actually have speech in them. And plenty of people just have a voice memo with 20 seconds of pocket noise before anyone starts talking.

A podcaster with a two-hour raw recording full of thinking pauses runs it through Shorten Long Pauses first to tighten the pacing, then boosts what's left with the Volume Booster & Normalizer so every segment sits at a consistent, comfortable loudness. Someone assembling three separately-recorded voice memos trims the dead air off each one here, then brings the cleaned clips into the Audio Merger to join them into a single continuous file — a smoother result than merging first and having to hunt for silence buried in the middle of the combined recording afterward.

What Automatic Silence Removal Can't Fix

Detection works on loudness alone, not on meaning — it can't tell the difference between a genuinely dead pause and someone speaking unusually softly, so a very quiet speaker can end up with parts of actual speech flagged as silence at an aggressive threshold. Previewing the red-shaded waveform before applying anything is the fastest way to catch this: if shading appears over something that clearly isn't silence when you listen back, the threshold needs loosening, not tightening.

It also isn't a noise-reduction tool — a recording with constant background hum or hiss during its "quiet" parts still has that noise there after silence removal, because the noise itself was never truly below the threshold in the first place. If a recording needs more precise, manual control over exactly where a cut happens rather than automatic threshold-based detection, the Audio Trimmer & Cutter handles that instead, with draggable handles for picking an exact range by hand.

Frequently Asked Questions

728 × 90 — Leaderboard Ad