Publishing and performance

    Brand kit

    Also called Brand-safe generation, Brand consistency.

    A brand kit is the saved set of colours, fonts, logo, product images, tone of voice and default frame shape that generations are expected to obey without being reminded.

    Generation is stochastic and a brand is the part that is not allowed to be. Every run is a fresh sample, so anything that must be identical across runs has to be supplied as a constraint rather than hoped for — which is the entire argument for storing it once instead of retyping it into every prompt.

    A saved kit and a habit of pasting the same paragraph are not the same thing. The kit is applied by default, so it survives the day somebody is in a hurry, the day a teammate runs the job instead of you, and the scheduled run that happens while nobody is watching. In Versely the kit is held per account and folded into later generations and workflows automatically once it exists.

    It is worth knowing where a kit binds hard and where it only leans. Overlay and caption styling is deterministic — a hex value is that hex value, a font is that font. Inside a generated image the same values are guidance, and the reliable way to hold a real product's shape, colour and label steady is to supply the product's own photographs as references rather than to describe them.

    In practice

    • Store the product shots, not just the colours — references constrain a generated product far harder than adjectives do.
    • Set the default frame shape in the kit so a vertical brand stops producing accidental 16:9.
    • Write the tone in the kit the way you would brief a writer: what it never says, not only what it sounds like.

    The mistake to avoid

    Treating a hex value in the kit as a guarantee inside generated imagery. It governs overlays and captions exactly; in a rendered scene it is a strong hint competing with the lighting the model chose.

    Where you will run into it

    Related terms

    The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.