What you say to the agent
No special syntax — just describe it like you would to a person.
What it does, step by step
- 1
You describe the subject, and the agent asks about style, aspect ratio, and how many images before generating — it won't guess and burn credits on the wrong look.
- 2
It picks a fitting model from Versely's image catalog (or uses the one you name), unless you're editing an existing image, in which case it routes to an edit-capable model automatically.
- 3
The image renders and lands in your generation history, ready to download, edit further, or drop into a slideshow or video.
What it needs from you
- •A text prompt describing what you want
- •Which AI model(s) to generate with (or let the agent pick)
What comes back
One or more generated images (your chosen count), saved to your generation history with the prompt attached.
What it costs
Priced per model you pick, shown before you confirm. Image models on Versely currently span a few credits to the high end for premium models.
Good to know
- The agent is instructed to ask about style, aspect ratio, and count before generating — expect a follow-up question rather than an instant render on a bare one-line prompt.
Under the hood
This is what the agent actually calls when you ask for it — real tools from its live surface, not marketing copy.
generate_imagesGenerate AI images from a text prompt. IMPORTANT: Ask the user about style, aspect ratio, and number of images before generating. Mention credit costs when suggesting models.
The full tool behind it
See it done in a real workflow
Morning Matcha Routine
UGC-style wellness reel — a clean-girl creator walks through her actual morning matcha ritual, then cuts to a cozy animated hero shot of the finished iced latte.
TrendingNYC Street Interview
Photorealistic street-vlog interview — Riley works three different NYC corners at golden hour asking strangers one question: "What's the wildest thing you've ever done?" Three candid OTS/two-shot clips with locked character references, real handheld energy and native spoken dialogue. Vertical 9:16.
TrendingPrimo Protein vs Other Brand
Pixar-style 3D comparison ad — the confident PRIMO PROTEIN pouch faces off against the tired old rival brand's tub across 11 talking clips: pasture vs dusty pantry, herb garden vs toxic lab, clean American lab vs grimy factory, chocolate-milkshake CTA vs "wet sand." ~40s vertical reel with native character voices.
Or start from a one-tap template
Frequently asked questions
What do I actually say to the agent to generate an image from text?+
Just describe it in plain English — for example: "Generate a photorealistic image of a cozy coffee shop window on a rainy morning" The agent handles picking the right tool and model from there.
What does the agent need from me first?+
At minimum: A text prompt describing what you want; Which AI model(s) to generate with (or let the agent pick). Anything else it needs, it asks for before running.
What do I get back?+
One or more generated images (your chosen count), saved to your generation history with the prompt attached.
Does this cost credits?+
Priced per model you pick, shown before you confirm. Image models on Versely currently span a few credits to the high end for premium models.
Anything I should know before asking for this?+
The agent is instructed to ask about style, aspect ratio, and count before generating — expect a follow-up question rather than an instant render on a bare one-line prompt.
You can also just ask for
Ask your Versely agent to generate an image from text
The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.