Guides

    Lighting Prompts for AI Video and Images

    Lighting prompts for AI video and images: the source-direction-quality formula, mood recipes you can reuse, and keeping light consistent across scenes.

    Versely Team7 min read

    Take any AI render that looks "off" in a way you can't name, and nine times out of ten the problem is light: it's coming from nowhere, it's coming from everywhere, or it disagrees with itself between the subject and the background. Generation models default to a flat, sourceless brightness because that's the statistical average of their training data — and average light is exactly what makes output look average. Lighting prompts are the highest-leverage sentence in your entire prompt, for video and images alike, because light is what your eye reads first, before subject, before style. Here's the formula and the recipes.

    Dramatic light falling across a scene

    The formula: source, direction, quality

    Every usable lighting prompt answers three questions in one clause:

    • Source — what is making the light? "Late afternoon sun," "a single desk lamp," "neon signage," "overcast sky," "firelight."
    • Direction — where is it relative to the subject? "From camera left," "backlit," "from below," "through a window behind her."
    • Quality — hard or soft, warm or cool? "Hard-edged shadows," "soft and diffused," "warm amber," "cold blue-white."

    Assembled: "lit by a single desk lamp from camera left, warm and soft, the rest of the room falling to darkness." That one sentence forces the model to commit to a physical lighting scenario — shadows fall coherently, the background darkens believably, and the render stops looking sourceless.

    The formula matters more than any individual buzzword. "Cinematic lighting" specifies none of the three questions and so changes almost nothing. "Golden hour" is better — it implies source (low sun), direction (raking), and quality (warm, soft) in two words — which is exactly why it's overused. Knowing why golden hour works lets you build a hundred alternatives.

    Six mood recipes you can reuse verbatim

    These are complete lighting clauses, tested across image and video models. Paste, then adjust the source to your scene:

    Mood Recipe
    Premium / editorial "soft window light from one side, gentle falloff, neutral white balance, deep soft shadows"
    Tense / thriller "single hard overhead light, harsh shadows under the eyes, cold tint, dark background"
    Cozy / intimate "warm practical lamps in frame, soft pools of light, everything outside the pools falling to warm darkness"
    Energetic / social "bright even daylight, slightly overexposed, minimal shadows, clean white background glow"
    Neon / nightlife "magenta and cyan neon signage as the only light sources, wet reflections, deep black shadows"
    Documentary / honest "flat overcast daylight, no dramatic shadows, true-to-life color"

    Notice each recipe names what the shadows do. Shadow behavior is the half of lighting most prompts forget, and it's where models betray you — a moody key light with cheerful bright shadows reads instantly wrong. Saying "background falls to darkness" or "minimal shadows" closes that loophole.

    Video-specific: light is a continuity contract

    In a single image, lighting only has to be beautiful. In video — and especially across multiple clips you'll cut together — lighting is continuity. An edit where shot one is golden hour and shot two is cool overcast reads as a time jump even if the action is continuous.

    Three practices keep a multi-shot AI sequence lit like one production:

    1. Freeze a lighting block. Write your scene's lighting clause once and paste it verbatim into every shot prompt for that scene. Don't paraphrase it — "warm window light" and "golden light through the window" will drift apart across generations.
    2. Change light only at scene boundaries. A new location or a time jump earns a new lighting block; a new camera angle doesn't. This mirrors how real productions relight.
    3. Let light carry the story's arc. The cheapest emotional structure in short-form: open in recipe one, end in recipe three. A skincare ad that moves from "bright even daylight" to "warm practical lamps" has told a morning-to-evening story without a word of copy.

    Lighting also compounds with camera choice — a slow push-in through hard side light produces drama neither achieves alone. Camera language is its own discipline, covered in the companion piece on camera movement prompts.

    Images vs video: where the same words behave differently

    The vocabulary transfers, but the tolerances don't.

    Image models reward maximal specificity — you can spec "hard rim light from behind-left, soft fill from a large window at right, warm practicals in the background" and a strong image model will honor most of it. Complex three-source setups are fair game when generating stills, and worth it for hero frames.

    Video models hold one or two light sources coherently across the clip; a third tends to swim — flickering, migrating, or fusing with another source mid-shot. Simplify to key-plus-ambience for video: one named source, one described fallback ("otherwise dim"). And for moving light itself — passing headlights, a flickering fire, a sweeping flashlight — name the movement explicitly ("firelight flickers across their faces"). Models render named light motion surprisingly well, but unprompted moving light is a common artifact you should treat as a defect and re-roll.

    Model tier matters here more than in almost any other prompt category: light coherence under motion is one of the clearest separators between premium and budget video models, and it's a major reason cinematic brand work clusters on the top tier — see the best AI models for cinematic brand films for that shortlist.

    Debugging light that won't behave

    • Everything looks flat. Your prompt has no direction word. Add "from camera left" or "backlit" — direction is the question models most often leave unanswered.
    • Subject and background disagree. You lit the subject but not the world. Extend the clause: "…and the same warm light catches the shelves behind her."
    • The mood is right but faces are ugly. Hard light is unflattering by physics, not by error. Keep the mood, soften only the face: "hard side light on the scene, her face caught in the softer bounce."
    • Light changes mid-clip (video). Reduce to one source and add "consistent lighting throughout." If it persists, it's a model-tier issue — spend up for the shots where it shows.
    • Color cast everywhere. Named colored sources bleed. Constrain the bleed: "neon glow on the wet street, skin tones stay natural."

    FAQ

    What's wrong with just prompting "cinematic lighting"?

    It answers none of the three questions a model needs — source, direction, quality — so the model substitutes its flat default. "Cinematic" describes an outcome, not a setup. Name a physical scenario instead: "single warm key light from the left, background falling dark."

    Why does golden hour work so well in AI prompts?

    Because two words encode a full setup: a low warm source, raking directional light, soft quality, and long shadows. It's a compressed lighting recipe. The lesson isn't to use golden hour everywhere — it's to prompt with terms that imply source, direction, and quality together.

    How do I keep lighting consistent across a multi-clip video?

    Write the lighting clause once and paste it verbatim into every prompt for that scene — paraphrasing invites drift. Only change the block at story boundaries like a new location or time of day, exactly where a real production would relight.

    Can AI video handle moving light sources?

    Yes, when you name the motion: flickering firelight, passing headlights, a sweeping flashlight beam all render well prompted. Unprompted light movement, though, is an artifact — if a static lamp starts wandering, simplify to one source and re-roll or step up a model tier.

    Should lighting prompts differ between image and video models?

    Same vocabulary, different complexity budget. Image models can honor multi-source setups; video models reliably hold one or two sources across a clip. For video, spec one key source plus an ambience fallback, and save the three-point setup for stills.

    Steal the six recipes, build your own lighting block, and browse more ready-to-run setups in Versely's prompt library — then A/B one scene flat versus lit and watch which one gets rewatched.