Silence removal
Remove silences from a recording automatically
Rescript cuts every pause of 0.3 seconds or longer out of a video or audio file in one click, on your own machine, for free. Each removal is an editable region on the timeline, so you can loosen a cut that's too tight or restore it entirely. No upload, no account, no subscription.
- Silence removal
- One click, pauses of 0.3s and over
- Filler word removal
- One click across the whole file
- Timeline
- Waveform, split, trim handles, draggable cut edges
- Price
- Free and open source for noncommercial use
- Where your media goes
- Nowhere. It never leaves your device
- Export formats
- MP4, WebM, M4A, MP3, WAV, SRT, VTT, JSON
Step by step
How it works, start to finish.
- 01
Open the file
Drop in a video or audio file. Whisper transcribes it locally with word-level timestamps, which is what lets Rescript tell a pause apart from a quiet word.
- 02
Click Remove silences
Every gap of 0.3 seconds or longer is cut across the whole file at once. The timeline fills with red cut regions showing exactly what went.
- 03
Tune the result
Drag either edge of any cut to leave a little more breathing room, double-click to reset it, or restore the pause entirely. Playback skips the cuts in real time, so you can listen to the new pacing before committing.
- 04
Export
Render MP4 or WebM up to 4K, or audio as M4A, MP3, or WAV. ffmpeg re-encodes on your own CPU and the file lands on your disk.
Silence removal that you can argue with
Automatic silence cutting is the highest-leverage edit in spoken video, and the one most likely to overreach. The beat before a punchline, the pause where someone is thinking, the breath that makes a sentence land — cut all of those and the result is tighter and worse.
So the automatic pass is a starting point, not a verdict. Every cut is a region you can see, drag, or delete, reinstating a pause takes a click, and the preview plays the edit as it stands.
It's one click of a larger edit
Removing dead air is rarely the only thing a recording needs. In the same session Rescript will strip every filler word, let you delete whole sentences by selecting them in the transcript, label who's speaking, and respeak a line you fumbled in that speaker's own voice.
Because all of it happens against one transcript with word-level timings, the cuts compose cleanly — remove silences, remove fillers, then delete the tangent in the middle, and the export is a single word-accurate render of what's left.
Nothing uploaded, nothing metered
Transcription runs through WebGPU on your own GPU, with a WASM fallback. Diarization runs locally. The export runs through ffmpeg compiled to WebAssembly on your CPU. Once the model files have downloaded once, the network is optional.
So there's no upload wait before you can start, no queue, no per-hour billing, and no limit on how long the recording can be beyond what your machine can hold.
FAQ
Questions people actually ask
How do I automatically remove silence from a video?
Open it in Rescript and click Remove silences. Every pause of 0.3 seconds or longer is cut across the whole file, each one shown as an editable region you can adjust or restore. It runs on your own machine and is free for noncommercial use.
Can I change the silence threshold?
The automatic pass uses 0.3 seconds. Individual cuts are fully adjustable afterwards — drag either edge to keep more of a pause, or restore it — and you can cut any additional gap by hand on the timeline.
Does it work for podcasts and audio-only files?
Yes. MP3, WAV, and M4A files get an audio-only layout, the same one-click silence and filler removal, and export back to M4A, MP3, or WAV.
Will removing silences desynchronise my video and audio?
No. Cuts apply to the media as a whole and the export is re-encoded with ffmpeg as a single word-accurate render, so picture and sound stay locked together.
Keep exploring features
Transcript editing
Delete a word, the footage goes with it.
Filler removal
Every "um" and "uh" in the file, in one click.
Speakers
Group the transcript by who is actually talking.
Timeline
Waveform, cut handles, and word-level timing by hand.
Regenerate
Rewrite a fumbled line, hear it in the original voice.
Weighing it against something else? See how Rescript compares.
Open it and see.
Free, open source, and running entirely on your own machine. No account, no upload, no watermark.
Or open the web app