Tighten a talking-head video (derush edit)
Goal: take a raw recording — a talking-head video, interview, vlog or screen recording with
narration — and get a tighter edit with long pauses, false starts and repeated takes removed. The
result is a new MP4 (edited.mp4).
"Derushing" is the editor's word for going through raw footage ("rushes") and keeping only the good parts. In ReelBolt an AI editor chooses which parts to keep, and ReelBolt itself does the precise cutting.
Prerequisites for a derush edit
-
The video uploaded to the project (Files tab). Up to 30 minutes and 2 GB per video.
-
A chat AI provider for the editor.
-
Strongly recommended: a transcription provider, so the editor knows what is being said and keeps whole sentences. An admin adds one under Admin, Inference Providers with capability
Transcription. Options:- the local
whisperservice that ships with ReelBolt (free, runs on your server; kindOpenAICompatible, endpointhttp://whisper:8000); - an Azure OpenAI or other OpenAI-compatible speech-to-text service.
Anthropic (Claude), Gemini and DeepSeek providers cannot transcribe. Without any transcription provider the edit still works, but the editor only sees pauses and shot changes, not words.
- the local
The derush workflow
Add the Highlight edit from footage template (video-derush-edit) from New workflow. Its steps:
| # | Label | Step type | Agent |
|---|---|---|---|
| 1 | Analyze source video | VideoAnalyze | VideoTransform (not an AI) |
| 2 | Decide which spans to keep | Agent | VideoStoryEditor |
| 3 | Compile edited video | VideoCompile | VideoTransform (not an AI) |
| 4 | Review edit quality | ReviewLoop | VideoReviewAgent (back to step 2 when below 8, up to 3 rounds) |
Step 1 — open it and select your video as the source. Its settings then look like:
{"version": 1, "source": {"kind": "ProjectFile", "projectFileId": "PUT-THE-FILE-ID-HERE"}}
Speech-to-text is used automatically when a provider exists ("transcription": "Optional"). Use
"transcription": "Required" to make the run fail clearly instead of editing without words, and
"language": "en" (or your language) if detection guesses wrong.
Step 3 — as shipped:
{"version": 1, "decision": {"from": "Previous"}, "analysisStepOrder": 1,
"transitionPolicy": "Auto", "programFadeInMs": 500, "programFadeOutMs": 800,
"programAudioFadeInMs": 300, "programAudioFadeOutMs": 900, "minSegmentMs": 800}
This makes soft automatic joins between cuts, fades the start and the end, and drops kept pieces shorter than 0.8 seconds.
How to run a derush edit
- Optional but useful: when you click Run, write a brief in the confirmation box to tell the editor what you want, for example "Cut to about 90 seconds, keep the story about the launch, drop the off-topic tangent near the end".
- Run it. Step 1 takes roughly as long as the video plus speech-to-text time; step 2 is one AI decision; step 3 encodes the video.
- Watch the result in step 3's Output Video, or in the project's Renders tab. The edit is also added to the project's Files.
What to check after a derush edit
- Step 1 result:
meta.transcription.appliedshould betrueif you expected words. If it saysdegraded, speech-to-text failed or no provider was found. If it showsdroppedNoSpeech, that many "words" were thrown away because there was no speech under them: speech-to-text tends to invent lines such as "Thank you for watching." on music-only or silent clips, and ReelBolt discards them so the editor never mistakes them for dialogue. Clips with no sound at all are not transcribed. - Step 2 result: the list of kept spans (which ids were kept) and the editor's explanation of why.
- Step 3 result: how long the result is, and the Open in Editor timeline to see each cut.
- Step 4 result: the review score and comments. A low score sends the run back to step 2 with the comments.
Variation: a panel of editors (edit room)
The Highlight edit, decided by an editing team template (video-derush-edit-room) replaces step 2 with an
EditRoom step: three AI editors (pacing, story, craft) discuss the footage and a director writes
the final decision. It is slower and uses more AI calls, but tends to make more considered choices
on long or story-driven footage. If the panel fails, a single editor decides instead
("fallbackToSoloEditor": true).
{"version": 1, "view": {"from": "Previous"}}
You can give the room your own seats, for example:
{"version": 1, "view": {"from": "Previous"},
"seats": [
{"name": "BrandEditor", "persona": "argues for keeping every product mention and the call to action"},
{"name": "PacingEditor", "persona": "argues for rhythm and removing dead air"}
]}
The room's discussion is kept on the step's result for reading; only the director's final decision counts.
Other variations of the derush edit
- Several takes or angles: in step 1 use
"sources": [ ... ]with several videos instead ofsource. All clips are analysed together and the editor can pick the best parts from any of them. - Open on the best shot, not the first one: by default the parts of one video play in the order
they were filmed. Set Order of the clips to In the editor's order in step 3
(
"keepOrder": "AsListed") and say what you want in the brief, for example "open on the moment the product launches, then tell how we got there". The editor may then move a later moment to the front. It never shows the same moment twice. This needs the normal (re-encoded) cut. - Use only part of a video: in step 1, fill in Use from and Use until (seconds) on a clip
to analyse and edit only that stretch of it (
"inSec": 12, "outSec": 48on the source). Everything outside it is left out of the edit. - Keep part of a long shot: set Split shots longer than in step 1 (for example 4 seconds,
"maxShotSec": 4). Long shots are then cut into shorter parts, at a calm moment when possible, and the editor can keep only the good part of a long shot. - Plain cuts: set
"transitionPolicy": "Off"in step 3 (and the fades to 0). - Fast, lossless cut:
"mode": "StreamCopy"with"allowKeyframeSnapping": truein step 3. Cuts land on the nearest keyframe (less precise) and no extras (graphics, music...) are possible. - Describe shots with AI:
"vision": "Optional"in step 1 adds a short AI description of what is in each shot (needs aVisionprovider; costs one AI call per described shot, at mostmaxCaptionedShots, default 50). - Add extras: graphics, music, sound effects, colour grade, b-roll and voiceover each have their own recipe in this section.
Pitfalls of a derush edit
- Run fails at step 1 with
SOURCE_UNRESOLVED: no video selected in step 1. EMPTY_KEEPorEXPECT_FAILEDat step 3: the editor kept nothing, or less than 15% of the footage. Give clearer instructions, or lowerexpect.minRetainedRatioin step 3 if a very short cut is intended.- Sentences cut mid-word: usually no transcription. Add a transcription provider.
- Re-runs give the same edit: the template sets the editor to never reuse, but if you built your
own workflow, set the editor step's
cacheModeto"Never". - If you add steps between the editor and the compile step, change step 3's
decisionfrom{"from": "Previous"}to{"from": "Step", "stepOrder": 2}— "Previous" would then point at the wrong step.