Natural-language editing
Type or dictate an editing instruction and Clip Smasher carries it out on your own footage.
A ten-minute piece to camera is usually seven minutes of talking and three of thinking. Clip Smasher will find the quiet parts and cut them in one instruction. This is what it does, what it cannot know, and how to tidy up after it.
It is loudness detection, not speech recognition. The app reads the audio levels across the clip, finds the stretches that are quiet relative to the rest of it, and cuts them out.
It cannot tell a voice from a motorbike, and it does not know what you said. Understanding that is the difference between it working brilliantly and surprising you exactly once.
Three rules keep the result listenable:
Import the clip and drag it onto a video track. Levels are measured when the file is imported, so the analysis is already waiting by the time you ask for it.
Open the command panel and type or dictate the instruction. All of these work:
The cuts are applied straight away, all together as one step. Play the whole thing through before doing anything else.
If the result is obviously wrong for the recording, undo takes every cut back in one press. That is usually a sign about the audio rather than about the app, and the next section covers what to do.
Loudness cannot tell the difference between a thinking pause and the beat before a punchline. Step through the cuts with the frame chevrons and nudge any that landed badly. A short crossfade on the audio hides a join that clicks.
Removing silence from a locked-off shot leaves visible jumps where you moved between takes. The usual fixes: cut b-roll over the joins on a second video track, or lean into it and let the jump cuts be the style.
| What you see | What is going on | What to do |
|---|---|---|
| It says there is nothing quiet enough | The recording is loud from end to end: usually constant background noise, air conditioning, traffic or wind. There is no contrast to find. | Apply noise reduction to the clip first, then run the silence removal again. |
| It says nothing is loud enough to keep | The whole recording is quiet. Often a microphone that was not connected, or a phone in a pocket. | Normalize or raise the clip volume first so there is a real difference between talking and not talking. |
| It kept the motorbike | Working as designed. It keeps loud things; it does not know which loud things are you. | Delete those pieces by hand afterwards. It is still far less work than cutting the silences yourself. |
| The speech sounds clipped | Rare, since short gaps are protected, but it happens on very fast delivery with hard consonants. | Undo, and instead trim the ends only with “top and tail every clip”, then cut the middle by hand. |
Often what you actually want is the dead air at the start and end gone and the middle left alone: the bit where you walked to the camera and the bit where you walked back to stop it. Ask for the edges specifically and the middle is untouched:
And once the piece is tight, the rest of the edit is the same sentence: instructions can be combined, so “chop out the parts where nobody is talking, add a crossfade between every clip, duck the music under the dialogue” is one command, one plan and one undo.
By loudness. It reads the audio levels of the clip and finds the stretches that are quiet compared with the rest of it, then cuts those out and keeps what is left. It is level detection, not speech recognition: it can tell loud from quiet, but it cannot tell a voice from a passing motorbike.
No. A gap has to last at least half a second to count as silence worth cutting, because removing the natural pauses between words makes speech sound clipped and unnatural. Short breaths and beats between sentences survive.
Then there is no contrast between loud and quiet and the app says so rather than cutting arbitrarily. Try noise reduction on the clip first, then run the silence removal again. If nothing in the recording is loud enough to keep, it tells you that too.
Yes. Ask for the edges only and it leaves the middle alone: “top and tail every clip”, “cut before the speaking starts”, “trim off the tail”.
Yes. The whole operation is applied as a single step, so one press of undo restores the clip exactly as it was.
Type or dictate an editing instruction and Clip Smasher carries it out on your own footage.
Video, audio, text and image tracks, trimming, splitting, stacking and frame-accurate navigation.
The twelve tape dials, what each one does to the picture, and three recipes to start from.