A city trip proves a wearable AI assistant
A live-action travel film turns three everyday questions into visible assistant demos before a clean product reveal.
Field notes / why it works
The camera appears on the traveler by 0:02.2 and answers a practical route question by 0:09, so the film proves its premise before explaining it. The scenes escalate from route and weather to visual answers and a real two-way translation exchange; the highest replay peak at 1:16.9–1:18.1 falls on that translation payoff. Seventy-three shots and recurring UI bubbles keep the 125-second film moving while a plain product card at 1:47 makes the offer clear.
Plate 01 / The format in one breath
A traveler moves through a city with a tiny wearable camera. Each ordinary travel friction—finding a route, identifying a landmark, speaking another language—becomes a short question, a visible product response, and a lifestyle payoff. The product name and offer arrive after the viewer has seen it work.
Why it works
- A moving tram and street whip open the film; the camera is on the traveler's chest by 0:02.2, before the first spoken question at 0:04.3.
- Each claim has a situated demonstration: route and weather by 0:29, photo-based landmark answer at 0:40, translation at 0:55–1:18. The UI bubbles make otherwise invisible speech legible.
- The strongest replay peak, 1:16.9–1:18.1, sits at the end of the translation exchange. This suggests (inferred) that the exchange is the most compelling demonstration, though a replay peak can also reflect viewers checking the wording.
Plate 02 / Format card
| Reference | |
|---|---|
| Platform · aspect · length | YouTube · 16:9 · 640×360 analysis copy · 29.97 fps · 125 s |
| Pace | 73 detected shots; 72 cuts; 34.6–49.0 possible cuts/min including uncertain motion changes; 37.9 confirmed picture changes/min; mean shot 1.71 s, median 1.07 s |
| Script | 164 transcribed words; 82 wpm over the whole film, about 153 wpm while talking; 23 short sentences averaging 7.1 words; 8 questions. The 1:21–1:41 automated transcript is unreliable. |
| Voice | Curious, conversational adult traveler; calm, concise synthetic-sounding assistant replies; a service worker in the translation scene; brief warm campaign narration (archetypes). |
| On screen | Predominantly live-action city b-roll and traveler action; wearable/handheld product close-ups, POV inserts, translucent response bubbles, occasional action-camera footage, white end card. |
| Captions | Small white sentence subtitles near bottom; contextual questions and answers in soft white/lilac rounded bubbles near the relevant object; minimal bold white feature labels. |
| Sound | Continuous music-like bed, about 9.5 dB below voice in the measurement; roughly 125 BPM (low-confidence estimate); street/activity sound and dialogue; −17.3 LUFS integrated, −2.8 dBFS true peak. No clear recurring cut SFX detected. |
| Structure | Six movements: city cold open → route/weather → visual answers → translation → travel-edit montage → product end card/CTA. Feature payoff at 1:13–1:18; product positioning at 1:41–1:54; subscribe/learn-more end screen 1:55–2:05. |
Pro / 13 plates + recipe
The first plates are open. The full dissection is Pro.
Get the timed shots, type and sound specs, and the runnable recipe.
Open the full blueprint- Plate 03 / The hook, frame by frame (first 3 seconds)Pro
- Plate 04 / Structure (beat sheet)Pro
- Plate 05 / Shot grammarPro
- Plate 06 / Visual style guidePro
- Plate 07 / Voice and deliveryPro
- Plate 08 / Sound designPro
- Plate 09 / PackagingPro
- Plate 10 / Script templatePro
- Plate 11 / Production planPro
- Plate 12 / PromptsPro
- Plate 13 / Edit recipePro
- Plate 14 / QA checklistPro
- Plate 15 / Make it yoursPro
- Recipe / runnable promptPro