Guides

    HeyGen Avatar V5: the mouth pass after the picture exists (4K, 50cr)

    HeyGen Avatar V5 wants a plate and a track. It is not how you invent the scene.

    Versely Team3 min read

    HeyGen Avatar V5 wants a plate and a track. It is not how you invent the scene.

    HeyGen Avatar V5 is HeyGen Avatar V digital twins — natural talking-avatar videos from a script or audio with premium lip-sync. Content type: lipsync. Category: premade-avatars. Audio: true. 4K. 50 credits. Requires image: false (the twins are premade). Durations: not listed as a ladder. The AI lipsync tool is the mouth pass. A text-to-video generate is the scene pass. Do not confuse them.

    The twin already exists. Your job is the line.

    Premade avatars means you are not casting a new face from a paragraph. You are putting words on a digital twin. Script or audio in, talking-avatar out, premium lip-sync. Features: Natural, Premade avatars, Script or audio.

    If you do not have the words, you do not have this job. Write the script. Record the track. Then open V5. Using V5 to "see what they might say" is how 50 credits buys a take you will rewrite.

    Lipsync is the category of work: mouth to track. It is not world-building. Background, wardrobe, and blocking are mostly already decided by the twin. If you needed a custom scene — raining street, product table, three-shot — that is a video generate plus a later mouth pass, or a different talking pipeline. V5 is the twin talking.

    Qualities: 1080p, 4k, 720p. Ceiling: 4K. Aspects: 16:9, 5:4, 1:1, 4:5, 9:16, auto. Pick the ratio the account ships. Auto is not a creative direction.

    50 credits is the mouth, not the movie

    Snapshot credits: 50. Billing is per second (10 credits per second on the matrix, min 50, max 1200). The title's 50cr is the floor, not a flat film budget. A long script is a long bill. Cut the line before you generate. A 30-second throat-clear is how min 50 becomes hundreds without a usable hook.

    Audio is true because the whole job is audio. No track, no lips. Do not feed silence and hope for a personality.

    Requires image is false on this row because the twins are catalog talent, not your JPEG. If your job is your face, this premade line is the wrong HeyGen (and the wrong expectation). Stay on premade when premade is the talent.

    After the picture exists

    "The mouth pass after the picture exists" is the claim. For V5 the "picture" is the twin you picked. Lock that choice the way you would lock a still: same avatar, same ratio, same 4K, every episode. Then swap only the script.

    If you do not want to be on camera, without showing your face is the broader pattern. Best AI model for talking heads is the ranked cut. This page is Avatar V5: premade twins, script or audio, 4K, 50-credit floor, lipsync, not a scene inventor.

    FAQ

    Can HeyGen Avatar V5 invent a custom scene from a prompt?

    That is not the job. Category is premade-avatars. Content type is lipsync. You pick a twin and give it a script or audio. World-building belongs on a video generate.

    Why is it 50 credits?

    50 is the snapshot credit figure and the matrix minimum. Billing is per second. Shorten the line. Do not treat 50 as a feature-length voucher.

    Do I need to upload my own photo?

    Requires image is false. These are premade digital twins. If you need your face, pick a row that takes a likeness, not this premade line.

    Does it include audio?

    Audio is true. The take is the speech: script or audio in, premium lip-sync out. No track, no job.