DEMO
A clip produced by ClipFlap, untouched
This page shows the captions themselves, on a clip left untouched: word-by-word karaoke aligned to the voice, line breaks that follow the phrasing, keywords highlighted automatically. The clip is a genuine ClipFlap output.
Extract with no audio track — the burned-in captions carry the speech.
Demo: word-by-word captions on a vertical clip
- Length
- 0:28
- What ClipFlap did
- Word-by-word captions aligned to the voice, line breaks that follow the phrasing, keywords highlighted automatically. 1080×1920 output, ready for TikTok, Reels and Shorts.
How this clip was produced
Timing: the start and end of every word come from the transcript's word timestamps, then a per-clip alignment pass checks them again against the audio to reduce drift between the captions and the voice.
Typography: the style is rendered by libass with the font of the language, and the current word is highlighted by a pill drawn on the real ink of the letters, not on the font cell.
Line breaking: lines are measured with the font's own metrics, French punctuation keeps its non-breaking space, and Arabic lines run right to left with connected letters. Nothing was edited after rendering.