Skip to content
refs
REF-1306Source onlineBlueprint rev 01

Four worlds, one listening promise

A 30-second music-led earbud film moves through four backstage and everyday worlds, then joins them with a two-part cyan message.

Maker @BoseFor BoseOriginal post ↗
Cut timeline · 21 shots00:30
00:0000:0500:1000:1500:2000:2500:30

Field notes / why it works

The earbud is visible in the first frame, and the 1.71-second cut opens its first social setting before returning to a close view. Twenty-one shots in 30 seconds move the same product through contrasting blue, cream, green and purple worlds; the first centred message at 20.5 seconds aligns with the strongest replay peak. Brief scene dialogue arrives only after 16.5 seconds, so the beat and images establish the premise first.

01

Plate 01 / The format in one breath

A music-led brand film moves a small product through several performers' workspaces. It opens on a close product-in-use detail, tours contrasting environments, then turns crew chatter and two brief text reveals into the promise that the product fits every moment. The product stays visible without stopping the film for a feature explanation.

Why it works

  • The first frame shows the white earbud against a face and saturated blue set. A cyan brand sweep crosses the close-up before the first cut at 1.71 seconds.
  • Twenty-one shots in 30 seconds, with eight cuts in the first ten, turn each new room into another use case while the earbud recurs at 0:00, 0:06–0:09 and 0:17–0:25.
  • The phrase “Hear it all” arrives at roughly 0:20.5 over a sudden foreground/background focus split; this is the strongest replay peak. “All the time” follows at 0:24–0:26 across two users.
02

Plate 02 / Format card

Platform · aspect · lengthYouTube · 16:9 · 640×360 capture at 23.98 fps · 30.0 s
Pace21 shots; 20 measured cuts, 40–44 cuts/min allowing two uncertain changes; mean 1.43 s, median 1.08 s, longest 4.49 s
Script29 recognised spoken words; first word at 16.48 s; 158 wpm over the speaking span, 186 wpm while someone speaks; transcription has eight low-confidence words
VoiceSeveral brief, natural on-set voices, rather than one narrator; speech covers about 37% of runtime
On screenLive-action portrait/product detail, backstage wide shots, home/workroom details, dark performance imagery, a camera-monitor view
CaptionsNo running subtitles; pale-cyan, centred keyword cards at 0:20.5 and 0:24–0:26
SoundContinuous prominent mid-tempo hip-hop/pop beat; incidental speech late; no clear systematic cut effects; −12.5 LUFS integrated, −1.8 dBFS true peak
StructureProduct hook → room relay → dark performance reset → on-set dialogue and message → monitor sign-off; no spoken CTA

Pro / 13 plates + recipe

The first plates are open. The full dissection is Pro.

Get the timed shots, type and sound specs, and the runnable recipe.

Open the full blueprint
  1. Plate 03 / The hook, frame by frame (first 3 seconds)Pro
  2. Plate 04 / Structure (beat sheet)Pro
  3. Plate 05 / Shot grammarPro
  4. Plate 06 / Visual style guidePro
  5. Plate 07 / Voice and deliveryPro
  6. Plate 08 / Sound designPro
  7. Plate 09 / PackagingPro
  8. Plate 10 / Script templatePro
  9. Plate 11 / Production planPro
  10. Plate 12 / PromptsPro
  11. Plate 13 / Edit recipePro
  12. Plate 14 / QA checklistPro
  13. Plate 15 / Make it yoursPro
  14. Recipe / runnable promptPro