Speech, voice and audio

    Audio extend

    Also called Extend music, Song continuation.

    Audio extend continues a previously generated music track from a chosen point, rather than writing a new song or time-stretching the file.

    You need the prior generate's identifier, not an arbitrary MP3. A continue-at time picks where the new section should pick up; prompt and style steer the added bars. The point is to keep the take you already liked instead of rolling a new song and hoping the mood matches.

    Video-extend continues picture from the last frame — a different medium, same idea of continuation. Text-to-music is the original write. Duration on a video model is a menu of clip lengths, not a way to stretch audio. Time-stretching a file you already have is an edit, not this job, and it will chip the transients.

    Keep the new prompt close to the original settings if you want a seam instead of a gear change. The editing surface is extend a music track; the agent job is extend a song.

    In practice

    • Look up the original generate and extend that record; an upload is the wrong input shape.
    • Pick a continue point mid-phrase rather than on a hard fade, so the model has motion to continue.
    • If the bed is right and only short, extend it; if the bed is wrong, write a new one.

    The mistake to avoid

    Feeding an uploaded MP3 into extend, or rolling a new song because the first fade was early. Extend needs the generated record; a new prompt is a different track.

    Go deeper

    Extending a song continues a generated track you already have

    extend_music needs a prior Suno audioId. A new generate_music prompt is a different job, and using it as an extend wastes credits.

    Where you will run into it

    Related terms

    The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.