Why film the hook here
Every other setting on this list gives you something to hide behind. The selfie vlog does not. There is no counter, no cup, no product to pick up while you find your point — the entire clip is your face and what you say in the first two seconds, which is why the prompts here are about expression rather than action.
That constraint is also why it is the most portable format in the set. A selfie vlog shot on a couch, in a hallway or in a car all read the same, so it is the one you can capture when the idea arrives rather than when the location is available. Most of the clips in this collection were built as pure reaction beats for exactly that reason: pick the face first, write the line to it.
It is also the format where a viewer's tolerance for a bad opening is lowest. In a kitchen you get a second of grace because something is happening. Here, if the first frame is you settling into position, adjusting the phone or drawing breath, the clip is already over.
- A face at arm's length fills a vertical frame completely, so there is nothing else on screen competing for the viewer's eye.
- It matches how the viewer is holding their own phone, which is the closest thing short-form has to eye contact.
- Emotion carries the whole clip, so a strong reaction can rescue a mediocre line — the reverse is almost never true.
What has to be in frame
Set this up once and every prompt below works without redressing the room between takes.
- A background with something in it — a corner, a lamp, a plant. A flat wall makes the shot feel like a document scan.
- Two to three feet of space behind you so the background is not against your head.
- One consistent light source in front of you, at or slightly above eye level.
- Nothing readable behind you: no whiteboards, posters, screens or paperwork.
- If it is the couch prompts, a cushion behind your back so you sit forward rather than sink.
- For the mirror prompts, a clean mirror and a wiped phone lens — both show at this distance.
Shooting notes for selfie vlog
Light
Whatever light you have, get in front of it and stay still relative to it. The failure mode here is not bad light, it is changing light — turning your head past a window mid-sentence swings the exposure and the clip visibly shifts.
Framing
Slightly above eye level, tilted down a few degrees, close enough that your chin and hairline nearly touch the safe area. Leave headroom for a caption bar. Too far back is the single most common mistake in this format.
Motion
Handheld, and let it drift. A tripod-locked selfie vlog reads as a broadcast; the small motion of a held arm is one of the format's authenticity cues. Just do not walk while you deliver the actual hook line.
Sound
You are close to the mic, which is the only real advantage of arm's length. Speak at conversation volume rather than projecting — the intimacy of the format collapses the moment you sound like you are performing to a room.
The two lines every prompt shares
Each scene below is the part that changes. These two lines are the same across the whole set, so they are written once here rather than repeated on every card — append both when you copy a prompt.
Base style line
Vertical phone UGC, authentic skin texture, natural light, slight handheld motion.
Clean plate line
Clean plate only: no on-screen text, no captions, no watermarks, no logos, no UI overlays.
One more exclusion for selfie vlog
Exclude anything readable in the background. At this framing the background is small but legible, and a document, screen or address on a parcel behind your shoulder is fully readable when someone pauses the clip.
4 hooks with a line to say
Each one is a scene, a suggested opening line and the finished clip it produced. Shoot it, or paste the prompt and generate it.
- 1
Mirror Happy Tears
A person records a mirror selfie in a bathroom, wiping away happy tears while smiling broadly.
Say this
“Didn't expect to cry today, but here we are. So much joy!”
crying-happymirror selfieFull prompt to copy+
A person records a mirror selfie in a bathroom, wiping away happy tears while smiling broadly. Vertical phone UGC, authentic skin texture, natural light, slight handheld motion. Clean plate only: no on-screen text, no captions, no watermarks, no logos, no UI overlays.
- 2
Outdoor Vlog Excitement
A young person films themselves in a vibrant outdoor setting, excitedly gesturing and talking directly to the camera.
Say this
“You guys HAVE to see this! I'm so hyped right now!”
excitedtalking-head reactionFull prompt to copy+
A young person films themselves in a vibrant outdoor setting, excitedly gesturing and talking directly to the camera. Vertical phone UGC, authentic skin texture, natural light, slight handheld motion. Clean plate only: no on-screen text, no captions, no watermarks, no logos, no UI overlays.
- 3
Selfie Laughing Fit
A person holds the phone in a selfie-vlog style, bursting into genuine, uncontrollable laughter, looking directly at the camera.
Say this
“You're not going to believe what just happened. I'm still laughing!”
laughingtalking-head reactionFull prompt to copy+
A person holds the phone in a selfie-vlog style, bursting into genuine, uncontrollable laughter, looking directly at the camera. Vertical phone UGC, authentic skin texture, natural light, slight handheld motion. Clean plate only: no on-screen text, no captions, no watermarks, no logos, no UI overlays.
- 4
Smug Walking Vlog
A person walks outside, holding the phone in a selfie-vlog style, giving a confident, slightly smug look to the camera as if they know a secret.
Say this
“Some people just get it, you know? 😉”
smugwalking vlogFull prompt to copy+
A person walks outside, holding the phone in a selfie-vlog style, giving a confident, slightly smug look to the camera as if they know a secret. Vertical phone UGC, authentic skin texture, natural light, slight handheld motion. Clean plate only: no on-screen text, no captions, no watermarks, no logos, no UI overlays.
7 reaction beats, no script attached
These come from the premade hook library inside the app. They store an expression and a scene rather than a line, which is the right way round when you already know what you want to say and need a face to say it over.
Animated
Talking directly to the camera with an animated friendly expression.
Confused
Confused expression, head tilted, looking quizzically at the camera.
Crying (M)
A man in his late twenties speaking quietly to the camera, holding back tears, vulnerable raw expression.
Crying Confession
Mascara slightly running, sincere vulnerable confession to the camera from the couch.
Excited
Bursting with excitement to the camera, hands flailing slightly, bedroom background.
Laughing
Throwing her head back laughing on the couch, warm lamp light.
Thoughtful
Mid-thought, speaking gently to the camera from a cozy couch.
What this set covers
The emotional and shot-format spread across the scripted selfie vlog hooks. Gaps here are real gaps — the set reports what the room actually produced rather than filling every cell.
Emotion
- Crying Happy
- 1
- Excited
- 1
- Laughing
- 1
- Smug
- 1
Format
- Talking Head Reaction
- 2
- Mirror Selfie
- 1
- Walking Vlog
- 1
A three-clip selfie run
All three shot in one position, one after the other. The variation comes from the delivery, not the setup — which is exactly what makes this format fast.
- 1
Cold open on the reaction
No greeting, no setup. Start on the expression — laughing, confused, stunned — and let the first word land after the face has already registered.
- 2
The sincere middle
Drop the energy. Same frame, quieter delivery, leaning in slightly. The contrast between clip one and clip two is doing more work than either clip does alone.
- 3
The turn
Back up to the energy of the first, but pointed forward rather than backward — the thing you actually want them to do. One sentence, then stop recording immediately.
What goes wrong in this setting
- Starting the clip before you start talking. Trim the first half-second, always.
- Holding the phone too far away. At full arm's length in a vertical frame you are a small figure in a large room.
- Delivering to the lens the entire time. One look away and back, at the right moment, is what stops it reading as a recital.
- Filming against a flat white wall, which removes the only depth cue the format has.
Workflows that fit these hooks
The lipstick review is the published workflow closest to a straight selfie-vlog product piece, with every scene prompt visible. The panel reel shows how a talking head gets cut against reaction footage, and the podcast build is the one to read if your hook is going to run longer than a single clip.
Plush Lipstick Influencer Review
Sara the influencer reviews the Plush Baby Pink Lipstick in 3 scenes: greeting, product review with features, and promo code CTA. Consistent voiceover narration throughout.
Viral Panel & Reaction Reel
A 4-scene viral hook video. Scene 1-2: A professional panel expert discusses her before/after results with a product, denying she's gatekeeping. Scene 3-4: An influencer films a selfie-POV reaction, holding the product and confirming the expert's claims.
Podcast with me
Host invites guests on his podcast and talks about the future and Artificial Intelligence.
Finishing the clip
A hook is the first two seconds, not the post. These are the steps between a raw selfie vlog clip and something you can publish.
Who this setting is for
Frequently asked questions
Why do most of the selfie prompts have no written hook line?+
Because most of them came from the premade reaction library, which stores a face and a scene rather than a script. That is the right way round for this format — you pick the expression that matches the point you are making, then write the line into it.
Is this format worth using if I have a good location available?+
Use the location. The selfie vlog is the format for when the idea is time-sensitive or the point is personal enough that a setting would dilute it. If you have a kitchen, a gym or a street available and the idea works there, it will almost always be a stronger clip.
How close is too close?+
When your forehead or chin is cropped, you have gone past it. Aim for hairline to just below the collarbone, and remember that platform UI eats the bottom fifth and the top of the frame, so a face centred perfectly in the raw file sits low once it is posted.
Should I look at the lens or the screen?+
The lens, if you can stand it. Looking at your own image on screen puts your eye-line a couple of centimetres off, and while nobody consciously notices, the clip reads as slightly evasive. Covering the preview with a finger for the take is a workable trick.
Can these be generated instead of filmed?+
Yes. The scene lines are written as generation prompts — combine with the base style line and the clean-plate line to get a clip in the same shape as the preview on each card.
Other settings
Outdoor street
Motion the viewer can feel. A moving background does a job no static room can — it makes a clip feel like it is going somewhere before you have said what.
Office desk
The only setting in this set built around a second screen. Over-the-shoulder framing turns a laptop into a reason to look — and the thing you are reacting to lives inside it.
Gym
The one setting where the state of your body is the hook. Out of breath, mid-wipe, still shaking — physical aftermath is evidence the viewer can see before you claim anything.
Generate these hooks in Versely
The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.