Title vs Verbal Hook
Same four engines. Two channels. Split the load; never repeat.
▪ Title hook (on-screen text)
Hard job: work with the sound OFF.
Carries: often a visual payoff (demo), but on talking-head, just the hook as text.
Size: 3-8 words, glanceable in ~1 sec.
Feels like: the billboard from the highway.
Carries: often a visual payoff (demo), but on talking-head, just the hook as text.
Size: 3-8 words, glanceable in ~1 sec.
Feels like: the billboard from the highway.
▪ Verbal hook (spoken)
Job: deepen the pull once they've stopped.
Carries: an ARGUMENT or STORY (claim, paradox, confession).
Size: one breath, ~5-12 words.
Feels like: the voice in your ear.
Carries: an ARGUMENT or STORY (claim, paradox, confession).
Size: one breath, ~5-12 words.
Feels like: the voice in your ear.
Rule of thumb (a lean, not a law)
Payoff you SHOW → title often leads.
Payoff you ARGUE / NARRATE → verbal often leads.
Exception: talking-head title hooks carry claims with no visual at all. The
one real rule is that the title must work muted.
The split rule
Divide the four engines where you can; the verbal must ADD, not just
echo. Title sets the gap (muted-readable); voice brings arousal + the
maybe. Some core overlap is fine, muted and sound-on are different audiences.
Use templates the right way
Strip any template to its engine, then build your own
skeleton in your voice, on a topic you know. Invent, don't fill in blanks.
The three leaks
Fail 1: RedundantTitle = verbal. One wasted. → Split them.
Fail 2: Mute-blindPull only in audio; dead first frame. → Put a title on screen.
Fail 3: Sound-dependent titleOn-screen text needs the audio to make sense. → Make it stand alone.
Why the split exists: ~85% of Facebook video watched on mute (Digiday 2016);
on-screen text lifts watch time ~12% (Verizon/Publicis 2023); ~71% decide in
the first 3 seconds, first frame first (TikTok, 2026). Directional, not gospel.