Captions

How placement, timing and emphasis are decided, and how to change any of them.

Captions are burned in, timed per word, and placed around your subject rather than at a fixed spot on the frame.

Placement

The editor knows where the face and body are in every frame, so the caption block lands in empty space instead of across someone's chin. When the subject moves, the captions move.

If there's genuinely nowhere clear, the block shrinks rather than covering the speaker.

Timing

Transcription is word-level, so each word lights on the syllable it belongs to. Cards break on natural phrase boundaries rather than a character count, which is why lines don't split mid-thought.

Emphasis

The word carrying the point can get its own treatment — a heavier face, a colour flip, a size step. You can name the words yourself, or let the editor choose from the sentence's own stress.

To change one word by hand, select it in the transcript panel and use Highlight.

Asking for changes

make the captions bigger and put them at the top

use the mono look, all caps, no emoji

highlight every product name in orange

two words at a time, tighter timing

Limits

Captions come from your audio. A mis-heard word is a transcription fix, not a rewrite — the editor won't put words in someone's mouth.