Talking Head Edit is for speech-led videos such as monologues, interviews, lessons, and video podcasts. AI can choose the strongest takes, remove filler and repeated ideas, tighten long pauses, and add captions, B-roll, motion graphics, or music when requested.
What do I need before writing the prompt?
- Import one or more video or audio files with clear dialogue, then decide whether to process the whole recording or only part of it.
- For multiple takes or cameras, identify the main picture and supporting angles. Use AI multicam sync from Chapter 3 first when the recordings need automatic alignment.
- Optionally prepare B-roll, logos, brand assets, a Design Style, or a preferred caption style.
- Use Selection Mode for clips, transcript text, or a time range, or plan to
@mention the target. Also decide the target length, ideas that must remain, and what AI may remove.
Prompts you can copy
Edit @interview. Keep the clearest and most natural takes, remove filler words, repeated sentences, and obvious restarts, and shorten pauses longer than 0.8 seconds without making the delivery feel rushed.
Clean up only the first 60 seconds of the talking-head footage on @V1. Remove empty pauses and repeated ideas, then add clean captions on key phrases. Do not add music.
Turn @interview into a 45-second vertical short. Open with the strongest line, preserve the complete point, use @B-roll to cover visible jump cuts, and add a lower third when the speaker first appears.
State the target length, ideas that must remain, what AI may remove, and whether you want captions, B-roll, motion graphics, or music.
Edit manually
- Place the footage on the timeline and open Transcript. If it is hidden, use Workspace → Transcript.
- Choose the track that contains the dialogue. Clicking text moves the playhead to that moment.
- Select unwanted words or sentences and press Backspace or Delete. ChatCut removes the matching spoken range from the timeline.
- In clip view, drag transcript segments to reorder the matching clips. Use the pause controls to shorten or restore pauses that exist in the source recording.
- Open the Captions menu in the playback controls to show captions and choose a style.
- Drag B-roll, motion graphics, sound effects, or music from My Assets, Library, or Templates onto the appropriate tracks.
Walk through one interview clip
Start with one interview recording already imported into My Assets. The goal is a short clip that preserves one complete point, with readable captions and a finished video file.
- Put the recording on the active timeline and identify the passage to keep. Open Transcript or reference the asset in the AI panel; include the selected passage or time range when you want only that part edited.
- Use the 45-second vertical-short prompt above as a starting point. Ask for 9:16, name the point that must remain, and specify use existing media only. To set the frame yourself, select Aspect Ratio → 9:16 below the Viewer; see portrait framing.
- Play the edited passage from its first word through its last word. Keep enough context to understand who is speaking and what the claim means. Check any removed pause or restart for natural delivery.
- Turn on captions, select the correct caption track, and review the wording, timing, and portrait placement in the Viewer. Visible captions are rendered into the video picture; use Export → Subtitles separately if you also need SRT/TXT. See Transcript and Captions when transcription or caption readiness blocks the next step.
- Adjust the clip or ask for one precise revision. Use Undo to recover an unwanted change, or save a Version before a larger pass.
- Select Export → Video for that timeline and follow the queue. Choose rendering and resolution using Web vs Desktop Export; keep the editor open for browser rendering. Find the finished file through Export Queue, Downloads, and Retries.
The timeline remains editable after export. Keep the original project when you need a different caption style or another delivery version.
How do I know the edit is ready?
| Check | What to verify |
|---|---|
| Meaning | The selected clip contains a complete point and enough context; it does not change the speaker’s claim. |
| Rhythm | Cuts and shortened pauses sound natural when played together. |
| Picture | A visible jump cut is acceptable or covered by media you provided. |
| Captions | The active track has the correct words, speaker timing, and readable layout. |
| Delivery | The queue shows completion and the exported file opens with the expected picture and sound. |
If an asset is still processing, resolve that before retrying an edit or export. If a render fails, inspect its queue row and the relevant workflow recovery guide rather than starting the same request repeatedly. For current charges and plan limits, use Credits Policy and the export guide.
What happens afterward?
Both AI edits and transcript edits change the current timeline without rewriting the source media. Review speech continuity, visible jump cuts, and the active caption track. Use Undo or save a Version before another major pass.
To localize a finished speech edit, see AI Video Translation. For a presenter you do not have to film, see AI Avatars / Digital Humans.
See also Transcript and Captions, Motion Graphics, and Versions / Undo / Redo.