Wan 2.6 Text to Image: lock this still before you buy motion (4K, 3cr)
Wan 2.6 Text to Image is the stills row. Animate after approval. Do not generate the pack inside the video model.
Wan 2.6 Text to Image is the stills row. Animate after approval. Do not generate the pack inside the video model.
Wan, 3 credits, text-to-image, max output 4K. Description: "Text-to-image generation with WAN V2.6." requires_image is false. audio does not apply. Aspect ratios listed: 1:1, 4:3, 3:2. Styles: Realistic, Artistic. Qualities: HD, 4K. The catalog also marks the row discounted (2 credits on the discounted step). You are buying a still you can hold up in a review. You are not buying a 5-second clip that happens to start with a pretty frame.
4K is the approval surface
Lock ratio on this row, not in a video prompt. 1:1 / 4:3 / 3:2 are the listed aspects. If the pack has to ship 9:16 later, that is a different still or a crop you accept before anyone spends a video credit. Video models will happily invent a new face at 9:16. That is the failure this page is for.
3 credits (2 on the discounted catalog step) is the stills bill. Wan 2.7 Text to Video is a different bill, a different identity sampler, and a different argument in review. Generating the pack "inside the video model" means you never got a still anyone signed. You got motion that now has to be reverse-engineered back into a hero frame.
Text to image is the door. Wan's roster is the family, including 2.7 video. Best AI image generator is the ranked stills shelf. This page is only 2.6 stills at 4K / 3 credits.
Approve the pack, then animate
The sequence is boring on purpose:
- Generate stills here. 4K. Realistic or Artistic. One ratio from the list.
- Pick keepers. Kill the almosts.
- Animate the keepers on an image-to-video row — Wan 2.7 Image to Video if you want to stay in-family, or the image-to-video tool if you do not.
Skipping step 2 is how a brand kit becomes "whatever the video model felt like this morning." Wan's stills tip is straightforward scene description plus a style keyword. That is the register. Do not write a camera move into a stills prompt and then act surprised that there is no motion.
Because this SKU does not require an image, it is a generate, not an edit. If you already have a locked still and need a delta, that is an edit row. Do not re-roll 2.6 and call the new still the same pack member.
What you are not buying
You are not buying duration. There is no 5s on this record. You are not buying native audio. You are not buying 9:16 unless you crop after, because 9:16 is not in the listed aspects. You are buying 4K pixels at 3 credits so a human can say yes.
If the yes never happens, do not "just animate it and see." Motion hides a weak still for one loop and then every thumbnail, pause frame, and ad preview shows the weakness. The cheap part of the pipeline is this row. Spend it. Then spend the video row on a frame that already survived a pause.
FAQ
Is Wan 2.6 Text to Image a video model?
No. Content type is image. Category is text-to-image. 4K, 3 credits, no durations. Animate after you approve the still. Wan 2.7 is the in-family video generate if the job really is a scene from text.
Which aspect ratios can I set here?
The catalog lists 1:1, 4:3, and 3:2. Pick one before you generate a pack. Do not assume a later video model will honour a ratio this row never offered.
Why not generate the hero inside image-to-video?
Because then the hero is a frame the model invented while it was busy doing motion. Reviewers pause. They see that frame. Generate it on purpose, at 4K, for 3 credits, and only then buy motion.
Does this row need a reference image?
No. requires_image is false. It is a text-to-image generate. If you need to edit a locked still instead, use an edit model. Re-rolling 2.6 is a new still, not a retouch.