2025 releases · AI audio

    Every AI audio model released in 2025

    7 models from 5 providers, dated 2 April 2025 to 16 December 2025. They did not arrive evenly — they landed in 2 bursts, and this page is those bursts in order.

    The 2025 audio timeline

    2 April – 23 September 2025

    3 models · 3 providers

    Chatterbox, MiniMax and Suno shipped 3 audio models between 2 April 2025 and 23 September 2025. Of the 3, 2 have a spec page and 1 is folded into a parent model's page as tier or mode variants. Cheapest complete job in the window: 1 credit on Suno Sounds V5.

    • Longest single clip the catalog had offered — 120s on Chatterbox TTS.
    • New names on the roster: MiniMax, Chatterbox, Suno.
    ModelBuilt byCredits per jobMax outputCapabilities
    MiniMax SpeechMiniMax's text-to-speech engine — clear multilingual delivery backed by one of the largest named-voice rosters in the catalog, so…MiniMax2 credits (headline rate)—Text to audio
    Chatterbox TTSChatterbox text-to-speech — natural, expressive narration with inline vocal tags like <laugh> and <sigh> woven directly into the…Chatterbox2 credits (headline rate)—Text to audio
    Suno Sounds V5variantSuno Sounds V5 generates high-quality sound effects and background music from text prompts with looping, tempo, and key controlsSuno1 credit—Text to audio

    1 October – 16 December 2025

    4 models · 3 providers

    Cartesia, Chatterbox and Inworld shipped 4 audio models between 1 October 2025 and 16 December 2025. Of the 4, 1 has a spec page, 2 are folded into a parent model's page as tier or mode variants and 1 has no page of its own.

    • First voice clone model in the catalog — Cartesia Sonic 3.
    • New names on the roster: Cartesia, Inworld.
    ModelBuilt byCredits per jobMax outputCapabilities
    Cartesia Sonic 3variantLatest and most capable Cartesia TTS model. Multilingual, expressive, and supports emotion control, speed tuning, and voice…Cartesia4 credits (headline rate)—Text to audio, Voice clone
    Cartesia Voice CloneClone any voice using Cartesia AI. Upload a short audio sample to instantly create a personalized voice for text-to-speech…Cartesia8 credits (headline rate)—Voice clone
    Inworld Voice CloneClone any voice using Inworld AI voice cloning. Upload audio samples to create a personalized voice for TTS generation.Inworld2 credits (headline rate)—Text to audio, Voice clone
    Chatterbox TTS TurbovariantThe faster Chatterbox voice pipeline — same natural delivery and inline vocal tags, tuned for quicker turnaround on longer…Chatterbox7 credits (headline rate)—Text to audio

    When 2025 was busy

    5 of the twelve months carried a audio release; the busiest window was 1 October – 16 December 2025, with 4.

    April 2025
    1
    May 2025
    1
    September 2025
    1
    October 2025
    3
    December 2025
    1

    Who shipped audio in 2025

    Cartesia (2), Chatterbox (2) and Inworld (1) led on volume. SKUs, not quality — four tiers of one model count four times.

    ProviderModelsSpec pagesFirstLatest
    Cartesia221 October 20251 October 2025
    Chatterbox211 May 202516 December 2025
    Inworld1017 October 202517 October 2025
    MiniMax112 April 20252 April 2025
    Suno1123 September 202523 September 2025

    What 2025 moved

    Firsts, measured against every audio model released before them. “New name on the roster” is the provider label in the catalog, not the company — one lab can hold several labels.

    Dates are each model’s released_at value, walked in order and grouped until a window held 3 or more. Credits are what one complete generation costs per the model’s own price matrix; where it states none, the headline rate is shown and the model sits out the cheapest-in-window line.

    Other release years

    Frequently asked questions

    How many AI audio models were released in 2025?+

    Versely's catalog carries 7 audio models with a 2025 release date, from 5 providers, arriving in 2 launch windows between 2 April 2025 and 16 December 2025.

    What was the biggest AI audio launch of 2025?+

    1 October – 16 December 2025, with 4 audio models from Cartesia, Chatterbox and Inworld. Of the 4, 1 has a spec page, 2 are folded into a parent model's page as tier or mode variants and 1 has no page of its own.

    Which company released the most AI audio models in 2025?+

    Cartesia, with 2 of the 7 audio models dated 2025 — first on 1 October 2025, most recently on 1 October 2025. Chatterbox shipped 2, Inworld shipped 1, MiniMax shipped 1.

    What changed in AI audio generation in 2025?+

    Measured against everything the catalog carried before it: Longest single clip the catalog had offered — 120s on Chatterbox TTS; New names on the roster: MiniMax, Chatterbox, Suno; First voice clone model in the catalog — Cartesia Sonic 3.

    Run any 2025 audio model in Versely

    The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync — in your browser or on your phone.