Captions
How placement, timing and emphasis are decided, and how to change any of them.
Captions are burned in, timed per word, and placed around your subject rather than at a fixed spot on the frame.
Placement
The editor knows where the face and body are in every frame, so the caption block lands in empty space instead of across someone's chin. When the subject moves, the captions move.
If there's genuinely nowhere clear, the block shrinks rather than covering the speaker.
Timing
Transcription is word-level, so each word lights on the syllable it belongs to. Cards break on natural phrase boundaries rather than a character count, which is why lines don't split mid-thought.
Emphasis
The word carrying the point can get its own treatment — a heavier face, a colour flip, a size step. You can name the words yourself, or let the editor choose from the sentence's own stress.
To change one word by hand, select it in the transcript panel and use Highlight.
Asking for changes
make the captions bigger and put them at the top
use the mono look, all caps, no emoji
highlight every product name in orange
two words at a time, tighter timing
Limits
Captions come from your audio. A mis-heard word is a transcription fix, not a rewrite — the editor won't put words in someone's mouth.