The board is the plan. Multi-scene generation is the execute step: each boarded shot becomes its own generate, then a stitch. Collapsing the board into one long prompt is how a model compresses five beats into a muddle.
First-last-frame is a two-still interpolation, which is a two-cell board with a motion job in between. Sora's storyboard row is a model that takes that plan as its input shape. A slideshow is a locked stack of stills that may never become motion. None of those are each other.
Write the board so each cell is one beat, one camera, one duration the catalog actually sells. Then pick a model per cell. The movie job on Versely is that unit.
In practice
- One cell, one shot. If the camera has to travel to a new setup, that is a new cell.
- Attach reference stills to cells that must keep identity; do not hope the next prompt remembers the last.
- Do not ask one 8-second generate to play five boarded shots.
The mistake to avoid
Treating the storyboard as decoration and prompting the whole film in one box. The board was the job shape; the single generate threw it away.
Where you will run into it
- AI Movie Maker — One prompt. A finished short film.
- Sora 2 Pro Storyboard — Sora 2 with advanced storyboard scene generation
Related terms
Multi-scene generation
Multi-scene generation is building a video as separately generated shots that are then stitched, rather than asking one clip to carry the whole story.
First-last frame
First-last frame generation takes two stills — where the clip starts and where it ends — and generates the motion that gets from one to the other.
Slideshow
A slideshow is a sequence of stills — generated or uploaded — assembled as a carousel or a short video, rather than a single generated clip.
Duration
Duration is how long a generated clip runs, chosen before generation from whatever lengths the model supports rather than trimmed afterwards.
Text-to-video
Text-to-video is generation from a written prompt alone — you describe a shot, the model invents every frame of it, and no image or footage goes in.
The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.