2025 releases · AI audio

    Every AI audio model released in 2025

    9 models from 5 providers, dated 2 April 2025 to 16 December 2025. They did not arrive evenly — they landed in 2 bursts, and this page is those bursts in order.

    The 2025 audio timeline

    2 April – 23 September 2025

    3 models · 3 providers

    Chatterbox, MiniMax and Suno shipped 3 audio models between 2 April 2025 and 23 September 2025. Of the 3, 1 is folded into a parent model's page as tier or mode variants and 2 have no page of their own. Cheapest complete job in the window: 2 credits on Suno Sounds V5.

    • Longest single clip the catalog had offered — 120s on Chatterbox TTS.
    • New names on the roster: MiniMax, Chatterbox, Suno.
    ModelBuilt byCredits per jobMax outputCapabilities
    MiniMax SpeechText-to-speech audio generationMiniMax1 credit (headline rate)Text to audio
    Chatterbox TTSHigh-quality text-to-speech with ChatterboxChatterbox3 credits (headline rate)Text to audio
    Suno Sounds V5variantSuno Sounds V5 generates high-quality sound effects and background music from text prompts with looping, tempo, and key controlsSuno2 creditsText to audio

    1 October – 16 December 2025

    6 models · 3 providers

    Cartesia, Chatterbox and Inworld shipped 6 audio models between 1 October 2025 and 16 December 2025. Of the 6, 2 have a spec page, 2 are folded into a parent model's page as tier or mode variants and 2 have no page of their own.

    • First voice clone model in the catalog — Cartesia Sonic 3.
    • New names on the roster: Cartesia, Inworld.
    ModelBuilt byCredits per jobMax outputCapabilities
    Cartesia Sonic 3variantLatest and most capable Cartesia TTS model. Multilingual, expressive, and supports emotion control, speed tuning, and voice…Cartesia4 credits (headline rate)Text to audio, Voice clone
    Cartesia Voice CloneClone any voice using Cartesia AI. Upload a short audio sample to instantly create a personalized voice for text-to-speech…Cartesia8 credits (headline rate)Voice clone
    Inworld TTSHigh-quality text-to-speech powered by Inworld AI. Supports a curated library of expressive voices with multilingual capability.Inworld3 credits (headline rate)Text to audio
    Inworld Voice ClonevariantClone any voice using Inworld AI voice cloning. Upload audio samples to create a personalized voice for TTS generation.Inworld5 credits (headline rate)Text to audio, Voice clone
    Chatterbox STS TurboChatterbox speech-to-speech turbo voice conversion modelChatterbox9 credits (headline rate)Audio to audio
    Chatterbox TTS TurboTurbo-speed text-to-speech with ChatterboxChatterbox8 credits (headline rate)Text to audio

    When 2025 was busy

    5 of the twelve months carried a audio release; the busiest window was 1 October – 16 December 2025, with 6.

    April 2025
    1
    May 2025
    1
    September 2025
    1
    October 2025
    4
    December 2025
    2

    Who shipped audio in 2025

    Chatterbox (3), Cartesia (2) and Inworld (2) led on volume. SKUs, not quality — four tiers of one model count four times.

    ProviderModelsSpec pagesFirstLatest
    Chatterbox301 May 202516 December 2025
    Cartesia221 October 20251 October 2025
    Inworld2117 October 202517 October 2025
    MiniMax102 April 20252 April 2025
    Suno1123 September 202523 September 2025

    What 2025 moved

    Firsts, measured against every audio model released before them. “New name on the roster” is the provider label in the catalog, not the company — one lab can hold several labels.

    Dates are each model’s released_at value, walked in order and grouped until a window held 3 or more. Credits are what one complete generation costs per the model’s own price matrix; where it states none, the headline rate is shown and the model sits out the cheapest-in-window line.

    Other release years

    Frequently asked questions

    How many AI audio models were released in 2025?+

    Versely's catalog carries 9 audio models with a 2025 release date, from 5 providers, arriving in 2 launch windows between 2 April 2025 and 16 December 2025.

    What was the biggest AI audio launch of 2025?+

    1 October – 16 December 2025, with 6 audio models from Cartesia, Chatterbox and Inworld. Of the 6, 2 have a spec page, 2 are folded into a parent model's page as tier or mode variants and 2 have no page of their own.

    Which company released the most AI audio models in 2025?+

    Chatterbox, with 3 of the 9 audio models dated 2025 — first on 1 May 2025, most recently on 16 December 2025. Cartesia shipped 2, Inworld shipped 2, MiniMax shipped 1.

    What changed in AI audio generation in 2025?+

    Measured against everything the catalog carried before it: Longest single clip the catalog had offered — 120s on Chatterbox TTS; New names on the roster: MiniMax, Chatterbox, Suno; First voice clone model in the catalog — Cartesia Sonic 3.

    Run any 2025 audio model in Versely

    The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.