Four focus modes
Choose motion, visual prominence, speech or a mixed score according to the story you want.
Advertisement
Choose a 15, 30 or 60 second target and a focus. Local analysis proposes moments; you—not an opaque auto-export—approve the exact ranges that become the final clip.
Result proof
Input
Latest local clip
Local pass
Scenes + motion + visual prominence + real VAD
Result
Reviewed MP4 or alpha WebM
How to verify it: Inspect each thumbnail and editable start/end time, confirm chronological order, then play the rendered result with sound and compare its duration against the selected target.
Editorial review
Scene boundaries, measured motion, visual prominence and real audio-energy VAD produce a chronological proposal for a 15, 30 or 60 second reel.
±5 seconds
Target-duration guard
Choose motion, visual prominence, speech or a mixed score according to the story you want.
Every proposed start and end remains visible and editable before rendering.
The accepted video and audio ranges use the same millisecond filter graph to avoid drift.
Know before you process: Automated scores do not understand narrative importance. Review the selected moments, especially in long interviews and quiet scenes.
Prioritize motion and strong scene changes, then remove moments that do not fit the event story.
Use Speech focus to surface active passages, then choose the quote boundaries yourself.
Use Visual prominence or Mixed focus to find high-contrast, centrally structured sections, then review whether they actually show the product.
Build a platform-length draft without giving up chronological range control.
No. Scene sampling, motion and visual-prominence scoring, audio-energy VAD, range review and final encoding run in the browser.
Motion compares sampled frames, Visual prominence measures luma structure and centre contrast, Speech uses adaptive RMS voice activity, and Mixed combines those three signals. It does not identify a person or understand story meaning.
Yes. Proposals are sorted chronologically and expose editable start/end times. Remove any range before rendering.
The selector targets that length with an allowed five-second tolerance so a useful scene is not split into an unusable microclip.
Accepted video and audio ranges are trimmed with the same millisecond boundaries and concatenated in the same order in one FFmpeg filter graph.
Motion, Visual prominence and Mixed remain available. Speech focus reports that no audio or speech was found and leaves the clip unchanged.
Yes. Known-alpha WebM stays alpha-capable WebM. When WebM alpha metadata is unknown, transparency preservation is enabled explicitly by default because an opaque sampled frame cannot rule out sparse or later-frame alpha. Known-opaque clips export to broadly compatible MP4.
Auto Subtitles
Create an editable caption track for the source or final reel.
Video Editor
Combine the reel with layers, captions, grading and privacy controls.
Video Compressor
Reduce delivery weight after the editorial cut is approved.
Auto Highlights Help
Read the signal definitions, review workflow and limitations.
No model download, no upload, and no blind auto-export.
Find local highlights →