Pick a depth. Each prompt opens in your AI pre-loaded with the lesson. Click a row to preview the prompt.
AI video models hallucinate the easy stuff: a glass shattering in slow motion, a neon-lit alley, a horse galloping on a beach. They fail the hard stuff: a character who appears in shot 1 matching shot 2, a hand turning a door handle without six fingers, a lipsynced line that matches the phoneme. The mental model for advanced filmmaking is: the tool is a beauty engine. You must supply story, continuity, performance, and rhythm. Understanding WHERE the tool's ceiling is saves you weeks of fruitless iteration.
Four concrete capability ceilings as of late 2025: (1) Character consistency — Veo 3, Kling 1.6, Runway Gen-4 all support reference images but degrade across >3 shots. (2) Lip sync — dedicated tools (Sync.so, LipDub, HeyGen) are better than the generalist video models. (3) Physical interaction — hands touching objects is still fragile; plan around it. (4) Long-duration shots — most models cap at 10 seconds; a 2-minute film = ~12-20 shots.
# Shot feasibility checklist (ask before prompting)
[ ] Is there a recurring character? If yes → plan for consistency tools (ref images, LoRAs)
[ ] Is there dialogue? If yes → plan for a dedicated lip-sync pass
[ ] Is there hand-object interaction? If yes → compose the shot to hide or obscure it
[ ] Is the shot longer than 10s? If yes → split into two shots connected by a cut
[ ] Does the shot require specific blocking (character moves from X to Y)? → expect 5+ retries
[ ] Is there a fast-motion element (explosion, water)? → choose a model known for motion (Sora, Kling)
[ ] Is it mostly ambience (landscape, detail)? → any model will produce something usable in 1–2 tries