Skip to content
refs
REF-2939Source onlineBlueprint rev 01

Expressive Avatar Launch Film

A one-minute Synthesia launch film that moves from a staged claim to a matched avatar comparison and emotion-led examples.

Maker @synthesiaIOFor SynthesiaOriginal post ↗
Cut timeline · 5 shots01:00
00:0000:1000:2000:3000:4000:50

Field notes / why it works

The claim begins changing on screen by 0:00.85 and reveals the category around 0:05, long before speech starts at 0:11.8. A script-entry glimpse at 0:19 and matched old/new avatar view at 0:24 turn the promise into visible proof; three emotion cards at 0:29–0:38 extend that proof in a repeatable visual system. The film has only four firm cuts, but 19 in-shot steps keep the minute active.

01

Plate 01 / The format in one breath

A widescreen product launch film first names a breakthrough in large type, establishes the old limitation, then demonstrates the new capability through a UI glimpse, a direct comparison, and a rapid gallery of expressive examples. The central promise appears as “the world’s first” at 0:01 and resolves into the product category by about 0:05.

Why it works

  • It makes the claim legible before anyone speaks: the opening title builds from 0:00.25 to roughly 0:05.5, while the first spoken line arrives at 0:11.8.
  • It supplies different kinds of proof: a script field around 0:19, an old/new split at 0:24, then emotion-led avatar cards from about 0:29–0:38.
  • Large words and people alternate. In the final 18 seconds, an audio/benefit graphic at 0:42–0:47 gives way to a reprise of the avatar gallery at 0:53 and a branded close around 0:57.
02

Plate 02 / Format card

Platform · aspect · lengthX · 16:9 · 1280×720 · 29.97 fps · 60.0 s
Pace5 measured shots, 4 firm cuts; 4–13 cuts/min including uncertain changes; 19 in-shot steps, 23 total picture changes/min; 12.0 s mean shot
Script37 detected spoken words; 46 wpm over the whole film, about 191 wpm in speaking stretches; short demonstration phrases rather than continuous narration
VoiceMultiple adult avatar-presenter archetypes; upbeat, articulate, expressive delivery. The pitch measurements combine different voices and should not be read as one speaker.
On screenPredominantly motion graphics and typography, with avatar panels, one text-entry UI glimpse, and one side-by-side comparison
CaptionsNo persistent subtitles; centered headline cards, large emotion words, and small comparison labels carry the information
SoundContinuous tonal, cinematic/electronic bed (music-like, inferred); roughly 83 BPM with low confidence; -8.0 LUFS integrated in the source
StructureClaim → previous limitation → input and contrast → emotion gallery → benefit → availability and logo. No explicit spoken CTA.

Pro / 13 plates + recipe

The first plates are open. The full dissection is Pro.

Get the timed shots, type and sound specs, and the runnable recipe.

Open the full blueprint
  1. Plate 03 / The hook, frame by frame (first 3 seconds)Pro
  2. Plate 04 / Structure (beat sheet)Pro
  3. Plate 05 / Shot grammarPro
  4. Plate 06 / Visual style guidePro
  5. Plate 07 / Voice and deliveryPro
  6. Plate 08 / Sound designPro
  7. Plate 09 / PackagingPro
  8. Plate 10 / Script templatePro
  9. Plate 11 / Production planPro
  10. Plate 12 / PromptsPro
  11. Plate 13 / Edit recipePro
  12. Plate 14 / QA checklistPro
  13. Plate 15 / Make it yoursPro
  14. Recipe / runnable promptPro