Create · Versely Agent

    Create an AI cover of a song

    Same song, new arrangement.

    What you say to the agent

    No special syntax — just describe it like you would to a person.

    Make an acoustic cover of this track
    Reinterpret this song as a synthwave version

    What it does, step by step

    1. 1

      Give the agent the URL of an existing track (your own upload, or a prior generation).

    2. 2

      Describe the new style or arrangement you want.

    3. 3

      It generates a Suno cover of the track in that new style.

    What it needs from you

    • An audio clip to work from
    • Which AI model to use

    What comes back

    A reinterpreted version of the source track.

    What it costs

    Priced per generation, shown before you confirm.

    Under the hood

    This is what the agent actually calls when you ask for it — real tools from its live surface, not marketing copy.

    cover_music

    Generate a Suno cover/reinterpretation of an existing audio track from its URL (new style/arrangement over the same song). Use when the user wants a cover version of a track they have a URL for (their own upload or a prior generation), not a from-scratch new song.

    The full tool behind it

    See it done in a real workflow

    Or start from a one-tap template

    Frequently asked questions

    What do I actually say to the agent to create an AI cover of a song?+

    Just describe it in plain English — for example: "Make an acoustic cover of this track" The agent handles picking the right tool and model from there.

    What does the agent need from me first?+

    At minimum: An audio clip to work from; Which AI model to use. Anything else it needs, it asks for before running.

    What do I get back?+

    A reinterpreted version of the source track.

    Does this cost credits?+

    Priced per generation, shown before you confirm.

    You can also just ask for

    Ask your Versely agent to create an AI cover of a song

    The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.