DEMO

A clip produced by ClipFlap, untouched

This page shows the captions themselves, on a clip left untouched: word-by-word karaoke aligned to the voice, line breaks that follow the phrasing, keywords highlighted automatically. The clip is a genuine ClipFlap output.

Extract with no audio track — the burned-in captions carry the speech.

Demo: word-by-word captions on a vertical clip

Length
0:28
What ClipFlap did
Word-by-word captions aligned to the voice, line breaks that follow the phrasing, keywords highlighted automatically. 1080×1920 output, ready for TikTok, Reels and Shorts.

How this clip was produced

Timing: the start and end of every word come from the transcript's word timestamps, then a per-clip alignment pass checks them again against the audio to reduce drift between the captions and the voice.

Typography: the style is rendered by libass with the font of the language, and the current word is highlighted by a pill drawn on the real ink of the letters, not on the font cell.

Line breaking: lines are measured with the font's own metrics, French punctuation keeps its non-breaking space, and Arabic lines run right to left with connected letters. Nothing was edited after rendering.

10 free minutes, no card.