Pulling a Toy From the TV: VFX Breakdown
A one-minute result-first illusion that reveals its practical rig and layered screen and book composites.
Field notes / why it works
The hand-to-screen illusion is readable by 0:02, and the real prop is shown by about 0:06, so the reveal follows the payoff quickly. A 22-second TV explanation stays active through at least five internal image changes, then a second clean-plate and masking problem refreshes the story around 0:37–0:50. The final repeated gesture and cut explain the last piece before the 60-second stop.
Plate 01 / The format in one breath
A first-person hand apparently pulls a tiny character out of a television and places it in a book. A brisk narrator immediately reverses the trick, revealing a hidden support, magnetic pickup, screen replacement, digital paint-out, page scan, masks, and a final cut. The format is a result-first process explainer: show an impossible action, then prove each layer with a physical or visual example.
Why it works
- The impossible action is legible without context: TV, hand, and target object share the frame from 0:00; the miniature clears the screen by about 0:02.
- Explanation starts around 0:06, before the viewer has to wait long; the real prop and its pickup mechanism appear by 0:09–0:15.
- The process keeps presenting fresh evidence: red annotation and screen comparisons around 0:19–0:35; book-page removal and masking around 0:38–0:56.
Plate 02 / Format card
| Reference specification | |
|---|---|
| Platform · aspect · length | TikTok · 9:16 · 720×1280 · 24 fps · 60.0 s |
| Pace | 9 measured shots, 8 definite cuts/min, 6.67 s mean shot; 15 measured picture changes/min including within-shot steps. 23 uncertain changes make a loose 8–31 cuts/min range. |
| Script | 193 spoken words · 195 wpm overall, 211 wpm during speech · 11 sentences averaging 17.5 words · no spoken questions. |
| Voice | Single animated, conversational, medium-register adult narrator; matter-of-fact maker explaining a trick. |
| On screen | First-person handheld live action: TV/hand illusion (~40%), toy and book evidence (~30%), compositing demonstrations and annotations (~30%). Shares are visual estimates (inferred). |
| Captions | No creator-added speech captions visible; printed words on the book are props, not subtitles. |
| Sound | Continuous quiet tonal bed detected under nearly constant narration (music classification inferred); some cut accents likely. −16.3 LUFS integrated, −2.5 dBFS true peak. |
| Structure | Result (0–6), practical method (6–15), TV composite (15–37), book composite (37–50), final sleight/cut (50–60). No explicit CTA. |
Tools you'll need: a camera shoot · After Effects for motion graphics · ElevenLabs or a recorded voice · Claude alone: partly
Frame by frame / measured
Stills taken at measured moments, each labelled with its time. Agents get the same sheets over MCP.
Pro / 13 plates + recipe
The first plates are open. The full dissection is Pro.
Get the timed shots, type and sound specs, and the runnable recipe.
Open the full blueprint- Plate 03 / The hook, frame by frame (first 3 seconds)Pro
- Plate 04 / Structure (beat sheet)Pro
- Plate 05 / Shot grammarPro
- Plate 06 / Visual style guidePro
- Plate 07 / Voice and deliveryPro
- Plate 08 / Sound designPro
- Plate 09 / PackagingPro
- Plate 10 / Script templatePro
- Plate 11 / Production planPro
- Plate 12 / PromptsPro
- Plate 13 / Edit recipePro
- Plate 14 / QA checklistPro
- Plate 15 / Make it yoursPro
- Recipe / runnable promptPro