How does the AI decide what a "best moment" is?
It reads the timestamped transcript and looks for segments with internal structure: a setup that resolves, a specific claim, a change in direction, or an answer that lands. Then it checks whether the segment still makes sense with everything around it removed, because that is the condition a clip is actually watched under. Segments that fail that test rank low even when they sounded lively in context.
Does it just look for loud parts or laughter?
No, and on an uploaded recording it could not if it wanted to. Audio-energy detection is cheap to build and it is why so many highlight tools return montages of people laughing at nothing. The picker for an uploaded file receives the transcript and nothing else, so loudness is not an input at that stage at all. Energy analysis does exist in our stack, but it runs on the live path, where a moment has to be chosen mid-broadcast before any finished transcript exists. Plenty of the highest-scoring finds are completely flat on a waveform.
How many moments will it find in an hour of footage?
It varies by how dense the recording is, and the number is not padded to hit a target. A focused hour-long interview usually yields somewhere in the high single digits to the mid teens. A rambling hour might yield three. If the material is thin, a short list is the accurate answer rather than a failure.
Do I get timestamps, or just the finished clips?
Both. Every moment shows its position in the source video alongside the rendered clip. If you would rather do the craft work in your own editor, you can ignore the exports entirely and use this purely as a search layer, jumping straight to the timecodes it returned.
Is there a cheap way to test the detection?
Yes. Access starts with a 3-day trial for $1, which exists precisely so you can check the moment selection against your own judgement before committing to a month. It takes recordings up to two hours, and after the three days it is $39 a month, cancellable in one click.
How is the 0-100 rank calculated?
It grades how strongly the clip opens, how efficiently it reaches its point, and where its peak falls within the runtime. It is calibrated for ranking clips against each other rather than predicting view counts. Used as a sort order it is reliable; used as a forecast it is not, and nobody honest would sell it as one.
Can it find moments in a video I did not make?
The tool accepts any URL it can reach, so mechanically yes. What you are allowed to publish afterwards depends on copyright and the rules of the platform you post to, and that judgement belongs to you rather than to us.
Does it work on a live stream?
Yes, and this is the part most alternatives cannot do. Connect a YouTube, Twitch or Kick channel and detection runs while the broadcast is happening, so moments surface during the stream instead of after you have downloaded and processed a huge VOD.
Is there a limit on how long the recording can be?
Two hours per upload on every plan, the $1 trial included. There is a second limit worth knowing about on the long end: the transcript budget is 48,000 characters, roughly fifty minutes of speech, and a longer recording is sampled evenly across its full runtime rather than read end to end. Coverage stays even from the first minute to the last, but the density drops, so a brief aside in a two-hour file can fall between samples. Sending it as two or three shorter uploads costs you nothing extra and raises that density. For a marathon stream, connecting the channel and letting live detection run is a better route than processing the whole VOD in one piece.
Will it find the moment I already have in mind?
Usually, and that is the test worth running first. Take a recording where you know exactly which segment is the good one, run it, and see where that segment lands in the ranking. If it comes back near the top, the detection matches your taste well enough to trust on recordings you have not reviewed.
What if it misses something obvious?
Two things usually cause it. Either the moment depends on context outside the clip, which detection intentionally penalises, or it is a community reference the model has no way to weight. Scan the lower-scoring finds before discarding them, since near-misses often sit at rank eight rather than being absent entirely.
Can I change the length of the moments it returns?
The cut length is chosen to fit the thought rather than a fixed duration, so a complete idea is not chopped to hit forty seconds. You can adjust the in and out points on any clip afterwards and re-render if you want it tighter or want an extra beat of reaction at the end.
Does it remove pauses and filler inside a moment?
Yes, automatically. Ums, false starts and silences are cut out of the segment, which usually recovers several seconds and noticeably tightens the pacing. A good moment with three seconds of dead air in the middle of it stops being a good moment.
Are the found moments captioned?
Every one, with word-by-word animated captions burned into the video in your choice of eleven styles. Since most short-form viewing happens with the sound off, an uncaptioned moment is effectively an undiscovered one.
Can I use this on Zoom or Teams recordings?
Yes, if you export the recording as a video file and upload it. Multi-person grid recordings are handled with the split and grid layouts so participants stay legible after the crop to vertical.
Does the finder work for non-English recordings?
Transcription covers major languages and detection runs on the transcript, so it functions beyond English. Quality is strongest in English and does vary with language and audio conditions. Running one representative recording through the $1 trial will tell you more than any claim on this page.
Can I search my past moments later?
Yes. Everything you process stays in your library with its score, title and source attached, so the clip you cut three months ago is findable when it suddenly becomes relevant again.
What aspect ratios do the moments export in?
Vertical 9:16 by default for Shorts, TikTok and Reels, plus 1:1 for square feeds and 16:9 if you want the moment as a landscape cut for a newsletter or a site embed.
Do the exported moments carry a watermark?
No. Exports are clean on the $1 trial and on every plan above it, with no badge of ours anywhere in the frame, and you can apply your own logo through the Brand Kit if you want your mark on everything you publish.
Can I post the moments straight to social?
Yes. Connect TikTok, Instagram and YouTube and either publish immediately or queue moments on a schedule, which is generally the better use of a batch of finds than dumping all of them out on one day.
Is there an API for moment detection?
Yes, the same detection engine is exposed through a developer API, and there is an MCP connector so Claude can run a search conversationally. Both are documented in the
developer docs if you want detection inside your own pipeline.
How long does a search take?
Usually a handful of minutes, stretching out for a two-hour source or when a lot of jobs are running at once. You are not required to watch it run — close the tab and the shortlist is there when you come back.
Will it work on a video with poor audio?
Detection depends entirely on the transcript, so bad audio degrades it in proportion. Heavy background noise, overlapping crosstalk and very quiet mics all cost accuracy. If the audio is bad enough that you struggle to follow it yourself, expect the shortlist to reflect that.
Is a moment finder useful if I already have an editor?
Often more useful, not less. Editors are expensive to point at raw footage because reviewing is the slow part of their day. Handing an editor a scored shortlist with timecodes moves their hours from searching to actually making the clips good.
What do you do with my footage once the search finishes?
It is used to produce your results and is not published anywhere by us. The clips and the moment list belong to you. The
privacy policy sets out how the files and data are handled.
Can I stop paying once the trial ends?
Yes, in a single click from your account, and we send an email before the trial converts so it never turns into a surprise charge. Whatever you have already paid for stays usable right up to the date it was due to renew.
What are the alternatives to an automated moment finder?
Scrubbing the timeline yourself, which is accurate and does not scale past the recordings you can face re-watching. Searching the transcript for phrases you remember, which only ever returns moments you had already half-recalled. Reading comments and chat replays, which is real audience signal but exists only for content people already watched. Or paying an editor or assistant to review the footage, which works well and turns into a per-hour cost that grows with every recording you make. Each is a reasonable answer depending on how much footage is piling up.
How is this different from a general clipping tool?
A clipping tool is judged on its output — the caption style, the crop, the export. A moment finder is judged on its choices. If you want to see the same engine framed around finished vertical output instead, the
YouTube Shorts maker is the same detection with a publishing emphasis.