Talk2Camera
Menu
en-GB
content creation

Say the hook, show the name

Your first three seconds decide whether there is a fourth. Spending them on your own name is the most expensive introduction in video.

A short-form video has about three seconds to justify a fourth. That is not a rule anyone wrote down; it is what the scroll does to attention. A feed is not an audience, it is an audition, and the judgement is instant and unsentimental. Which makes the standard opening — "Hi, I’m Maya, welcome back to my channel, today I want to talk about…" — the most expensive sentence in the format. By the time it ends, the decision has already been made, and it was made by people who do not know you, on the only stretch of the video where their attention came free. The cruel part is that the sentence is not even working for the people it names you to. A name means nothing before a claim has earned it; nobody stays for a stranger’s name. They stay for the sentence that made a promise. Introducing yourself first is answering a question nobody has asked yet, and paying for the answer with the seconds that decide everything else.

Television solved this problem so long ago that the solution has a name older than any platform: the lower third. Watch a news broadcast and notice what the anchor never does — introduce themselves. The correspondent appears mid-sentence, already reporting, and their name, title and city sit in a strip along the bottom of the frame. The voice carries the story; the strip carries the identity. These are two different channels, and a viewer uses both at once without effort, because reading a short name takes nothing away from listening to a sentence. That is the actual insight, and it is about routing, not decoration. Identity is visual information; the hook is spoken information. Say the hook, show the name. A talking video that opens on its strongest claim, while a small card quietly answers "who is this?", does both jobs inside the same three seconds — and the viewer decides with full information instead of a greeting. Broadcast did not adopt this because it looked professional. It adopted it because a newsreader who introduced every story with their own name would be unwatchable, and the feed has only made that arithmetic harsher.

The conventions do need translating, because broadcast designed them for a television across a living room and a phone is a small screen at arm’s length, usually with the sound of a captioned world around it. Size first: text that looks respectable in a desktop edit becomes lint on a phone, so set the name larger than instinct suggests and test it the only honest way — on your own phone, held at arm’s length, without squinting. Contrast second: thin white lettering over moving footage survives exactly until something light passes behind it, which is why the convention is a backing — a solid or semi-opaque bar, or type heavy enough to hold its edge against any frame that might end up behind it. Timing third: a card that sits on the video from the first frame reads as a template, and one that flashes past cannot be read at all. The working rhythm is that it arrives just after the video starts, holds long enough to be read twice, and leaves before it becomes wallpaper. And position: the lower third of the frame, but clear of the strip along the bottom that every platform reserves for captions, progress bars and buttons.

This is also a job an export step does better than a person. Talk2Camera burns a name card — your name, and a handle if you want one — over the opening seconds of an export: you choose how long it holds, anywhere from 2 to 8 seconds, and it fades in and out gently instead of popping. Set it once and it is remembered across takes, so every video that follows carries the same card without you rebuilding it. Like captions, it is applied at export, which means the recording itself stays untouched — change your handle next month and nothing about the original take has to be redone. And it is positioned to stay clear of the corner mark that the free tier adds, so the two never collide. The same export pass that keeps word-highlight captions inside each platform’s safe margins handles the card by the same logic: the platform presets know where a destination’s interface lives, and the card stays out of its way.

What remains is the pleasant problem of spending the three seconds you got back. Open on the claim, the question, the result — the sentence you would previously have buried at second six, delivered to the camera as the first thing a stranger hears. The card answers the identity question at the exact moment it actually arises, which is a beat or two in, once the promise has landed and a viewer wonders who is making it. Returning viewers read nothing and lose nothing; new viewers get your name at zero cost to the hook. There is a quiet confidence in this arrangement that speaking your name cannot buy. An on-screen introduction says: the content will vouch for me, and here is my name when you want it. Your name matters — that is the point. It matters enough not to spend your best seconds saying it out loud to people who were never given a reason to remember it.

Mentioned in this article

Lower third

A name and handle card over the opening seconds — two to eight seconds on screen with a gentle fade, remembered across takes, burned in at export like captions. The recording stays untouched.

Word-by-word caption highlight

Burned-in captions can pop word by word as they’re said — the style short-form feeds expect — timed inside each line’s real window, with the same styles and presets.

Platform export presets

TikTok, Instagram, YouTube and LinkedIn, each with its own resolution cap, bitrate and duration guidance. The vertical feeds get a 9:16 frame; the others keep your shape.