Prompt pairs that cancel each other out
Contradictions like 'static handheld' make a model pick one instruction at random, take to take. The common conflicts and how to collapse each into one.
Contradictions like 'static handheld' make a model pick one instruction at random, take to take. The common conflicts and how to collapse each into one.
Shallow and deep focus are executable prompt instructions. A focus pull is not, unless you stage it across two generations. How to specify the focus plane.
Which camera moves survive as text, which get silently approximated, and what a motion-reference clip costs you in framing freedom. With a three-bucket rule.
A dolly and a zoom both read as getting closer, so models pick at random. How to name the physical mechanism and read the background to check what you got.
24mm, 85mm, macro and anamorphic shift field of view and depth of field, not just mood. Which lens words are load-bearing and which are decoration.
Leaving camera language out of a video prompt does not give you a static frame. The explicit phrasing and staging that produce a truly motionless camera.
Mixed edits break on grain, grade, focal length and motion blur. The order to fix them in, and the one thing you cannot repair in post.
Slow dolly is re-estimated on every generation. Parameter-based camera control and endpoint framing are the two ways to lock one move across a whole sequence.
Models default to mid-scale unless something anchors the frame. The lens, camera height and foreground cues that make a subject read enormous.
Leading a video prompt with shot size and angle changes the framing you get back. A five-slot shot-card skeleton you can fill for any shot, in order.
Some camera directives are real parameters. Others are prose the model approximates and quietly drops. How to tell which is which, and test for it.
Why whip pans and crash zooms break text-to-video prompts, what Wan 2.2 docs reveal about how camera motion is learned, and when to switch to motion transfer.
Kling O3 prompting guide: camera grammar that lands, cause-and-effect motion cues, image-to-video direction, and fixes for the classic failure modes.