Your country

Tools that support it use your country for local currency, number formats, units and paper size. Your choice is saved only in this browser.

Type a name or a two-letter code. Use the up and down arrow keys to move through the countries, Enter to choose one and Escape to close.

Remove Retakes (Cut Repeated Takes from Talking Videos)

Said a line three times? Keep the take that worked and cut the rest, after a review.

Audio & Video No upload Free preview, no sign-upIncluded in your pass Pro tool Pro pass: ₹179 for 30 days

Try before you buy.

  • Free preview: the first half of the result (up to 10 seconds) at the settings you chose, with a MySmartCoPilot watermark.
  • Locked until you unlock it: download and copy.
  • Unlock: Pro pass, ₹179 for 30 days, a one-time payment that never renews.

Ways to unlock shows how to get the full result.

See passes (opens in a new tab)

Printing this result is locked in the free preview.

Next steps

About the Remove Retakes (Cut Repeated Takes from Talking Videos)

Recording a talking-head video, a course, a voice-over or a podcast usually means saying a line again until it comes out right — and then scrubbing through the recording to cut the attempts you did not want. This page finds those retakes for you: it transcribes what you said on your own device (or reads the subtitles you already have), finds the sentences you said more than once within a minute or so, and proposes to keep the last take — or the cleanest, with the fewest “um”s, repeats and cut-off endings — and to cut the others, together with the “sorry, again” in between.

Nothing is cut until you have checked it. Every retake is listed with each take’s words and time and a Play button, so you can listen, choose another take to keep, or untick a retake to keep it all. The cuts go in the quietest moment of each pause, with a short cross-fade in the sound, and the result is a new MP4 (or MP3, M4A or WAV) plus the transcript of what is left, as text or SRT subtitles.

The free preview on this page is made of the parts around the cuts — up to 10 seconds of the video with a MySmartCoPilot watermark, or up to 30 seconds of the sound with a short chime — and the first words of the transcript. With a Pro pass, or after unlocking this one result, you get the whole file and the full transcript.

How to use it

  1. Choose or drop your video or recording (MP4, MOV, WebM, MKV, MP3, M4A, WAV, OGG or FLAC). It stays on your device.
  2. Press Transcribe on this device — pick the quality and the language first; the first time, agree to the one-time download of the speech-recognition model — or choose the SRT or WebVTT subtitles you already have for this file.
  3. Check the Retakes found: play the takes, pick the one to keep, and untick any retake you want to keep whole. Match (strict to loose), the time window and Keep (last or cleanest) change what is found; tick Also match takes that say the same thing in other words to compare sentences by meaning with an on-device model.
  4. Choose Save as and press Remove the retakes. Without a pass you get the free preview of the parts around the cuts; with a Pro pass, or once this result is unlocked, Download gives the whole file, and Copy transcript, Subtitles (SRT) and Text (TXT) give the transcript of the result.

Examples

A product review recorded in one go
Input
12-minute MP4 from a phone · the opening line said three times, two sentences said twice · keep the last take
Result
review-retakes-removed.mp4, about a minute shorter, starting with the take that worked, plus its transcript as SRT subtitles that match the new timing.
A voice-over for a course lesson
Input
WAV narration · false starts like “So in this lesson we— so in this lesson we’ll…” · keep the cleanest take · subtitles from the script
Result
lesson-retakes-removed.wav (lossless) without the false starts or the “sorry, again” between takes, ready for the video editor.

Common uses

  • Talking-head videos for YouTube, Reels and Shorts recorded without stopping the camera
  • Online course lessons, tutorials and screen-recorded explanations
  • Podcast and interview segments where a question or an answer was asked again
  • Voice-overs and narrations read from a script or a teleprompter

How retakes are found

  • What is said. The speech-recognition model runs in your browser and gives the recording’s sentences with their times; subtitles you bring work the same way, without AI. Long stretches are split at sentence ends, and a sentence that starts again part of the way in (“so today we, um, so today we’re going…”) is split where it starts again.
  • Takes of the same line. A sentence is compared with the four before it, as long as only short asides (“sorry”, “one more time”, “let me try that again”) sit between them and it comes within the time window (1 minute unless you change it). Two sentences are takes of one line when they share most of their words (word overlap and spelling similarity, set by Match), when one is the cut-off start of the other (a false start), or — with the sentence model on — when they mean nearly the same and still share some words.
  • The take kept. The last take by default, because people usually repeat a line until it is right; or the cleanest one, with the fewest fillers, words said twice and unfinished endings.
  • The cut. From the pause before the first take that goes to the pause before the take that stays, at the quietest 10 ms the sound has there, so no word is clipped; the sound is cross-faded over 30 ms at each join.

Languages, accuracy and checking

Speech recognition covers about a hundred languages, Hindi among them, and subtitles you bring can be in any language or script; matching by meaning uses an English model for English and a multilingual model for everything else. Like any speech recognition it can mishear names, accents, music and people talking over each other, and it sometimes leaves out a false start altogether — so treat the list as suggestions: listen to the takes before you save, and untick anything that is not a retake.

Limitations

  • Only retakes said as words are found: a take you repeated because of a cough, a noise or a gesture is not, unless the words differ.
  • Cuts follow the sentences of the transcript; a slip inside a long sentence that was never restarted is not cut.
  • Speech recognition can miss very short false starts and repeated words, which it tends to tidy away.
  • A video is encoded again (an H.264 MP4 at high quality, or WebM where the browser cannot write H.264) so the cuts can fall on any frame. Sound-only results keep the original’s channels: WAV at its bit depth (16 or 24-bit), MP3 or M4A at about its bitrate.
  • Up to 90 minutes on a computer and 20 minutes on a phone; cut longer recordings into parts first. Phones take noticeably longer than computers.

Privacy

Everything happens in your browser. What you enter or open here is not uploaded or stored by MySmartCoPilot. The video engine downloads from MySmartCoPilot the first time you open a file. To transcribe on your device, a speech-recognition model (about 44, 80 or 252 MB, by the quality you choose) downloads once after you agree; matching takes by meaning downloads a sentence model (about 24 MB for English, 135 MB for other languages) the same way. Your videos, recordings, subtitles and transcripts are never uploaded.

Frequently asked questions

What do I get without a pass?

Without a pass, Remove Retakes (Cut Repeated Takes from Talking Videos) shows the first half of the result (up to 10 seconds) at the settings you chose, with a MySmartCoPilot watermark. Until you unlock it, the result can’t be downloaded or copied. A Pro, Premium or Ultimate pass, a one-time payment that never renews, unlocks the full result. The pricing page lists the passes and their prices.

Is my video uploaded?

No. The file is read, transcribed, cut and written again in your browser on your own device. The AI models download to your device once, after you agree; your video, its sound and its transcript are never sent anywhere.

Does it cut anything without asking?

No. It lists every retake it found, with the take it would keep, and you can play each take, keep another one or untick the retake. Only when you press Remove the retakes is a new file made — your original is never changed.

Which take does it keep?

The last take of each line by default, since people usually say a line again until it is right. Choose The cleanest take to keep the one with the fewest “um”s, words said twice and cut-off endings instead — or pick any take yourself.

Can I use subtitles instead of the AI?

Yes. If you already have SRT or WebVTT subtitles for the file — from YouTube Studio, your editor or a caption tool — choose them under Or use subtitles you already have. Then no AI model is downloaded or run at all.

Does it work for Hindi and Hinglish?

Yes, with care. Choose the language spoken before you transcribe; speech that mixes Hindi and English is recognised less reliably than either language on its own, so check the takes it lists. Subtitles you bring can be in any language or script, Hinglish ones included. Matching by words works in any language; for matching by meaning the page uses a multilingual sentence model for every language other than English.

What does the free preview include?

Without a pass the page makes only the parts of the result around the cuts — up to 10 seconds of the video with a MySmartCoPilot watermark, or up to 30 seconds of the sound with a short chime — and shows the first words of the transcript, so you can check that the cuts land well. The whole file and the full transcript need a Pro pass, or an unlock of this one result, which makes it from the file still open on the page.

Quick answers and tool search

Type to search tools or to get a quick answer, for example 18% of 2500. Use the up and down arrow keys to move through the results, Enter to choose, and Escape to close.