English日本語한국어

Promptsugc-talking-headTokyo Golden Hour Selfie Vlog with Iced Latte Café Stop

Tokyo Golden Hour Selfie Vlog with Iced Latte Café Stop

A 15-second vertical smartphone-style vlog following a young woman walking through a quiet Tokyo street at golden hour, speaking to the front camera. She enters a cozy café, flips the camera to show an iced latte being made, then returns to selfie mode to sip and continue walking. The prompt emphasizes photorealistic front-camera quality, natural handheld motion, ambient city audio, and authentic influencer aesthetics.

Prompt (original, unmodified)

Create an ultra-realistic 15-second vertical (9:16) smartphone vlog. A young woman is filming herself with the front camera while walking through a quiet Tokyo street during golden hour. The camera has natural handheld movement, slight autofocus shifts, realistic lighting, and subtle background blur. People walk by naturally, bicycles pass in the background, birds chirp, and distant traffic creates authentic city ambience.

She smiles warmly at the camera and speaks naturally in English. She walks into a cozy café, briefly flips the camera to show the coffee machine and barista making an iced latte, then switches back to selfie mode. She picks up the drink, walks outside, takes a sip, and continues walking with the city behind her. The video should feel like a genuine influencer vlog, not scripted or cinematic.

Voice & Timing:

0–4 sec:
"Good morning, everyone! I found the cutest little coffee shop while exploring Tokyo."

4–9 sec:
"Let's see if their iced latte is really as good as everyone says."

9–15 sec:
(Takes a sip and smiles)
"Wow... that's amazing. Definitely worth the stop! See you in the next adventure!"

Style & Audio:

Ultra-photorealistic, 4K HDR

Natural facial expressions and lip sync

Realistic smartphone front-camera quality

Ambient city sounds only (birds, footsteps, distant traffic, café sounds)

No background music

Authentic influencer vlog aesthetic

Smooth transitions, realistic motion blur, and no AI-looking artifacts

Reproduced verbatim from the original post. Never edited or translated. Everything outside this block is our own commentary.

How this prompt compares

Computed from the 5,926 prompts in this library, not from this prompt alone. How the data is produced.

The techniques this prompt uses are more common in UGC prompts than in the library overall:

TechniqueIn UGCLibrary averageLift
handheld camera78%24%3.22x
close-up camera64%40%1.6x
natural light lighting64%28%2.26x

"Lift" is how much more often the technique appears in UGC prompts than across the whole library.

Across all 5,926 prompts, the techniques used here appear at: close-up 40%, natural light 29%, handheld 24%, golden hour 17%.

The 15s duration is not arbitrary: 71% of the prompts in this library that state a duration ask for exactly 15 seconds — Seedance 2.0 accepts clips up to 15 seconds, so 15s is the model's ceiling and the most common choice.

Related cases