The 2025 audio timeline
2 April – 23 September 2025
3 models · 3 providersChatterbox, MiniMax and Suno shipped 3 audio models between 2 April 2025 and 23 September 2025. Of the 3, 1 is folded into a parent model's page as tier or mode variants and 2 have no page of their own. Cheapest complete job in the window: 2 credits on Suno Sounds V5.
- Longest single clip the catalog had offered — 120s on Chatterbox TTS.
- New names on the roster: MiniMax, Chatterbox, Suno.
| Model | Built by | Credits per job | Max output | Capabilities |
|---|---|---|---|---|
| MiniMax SpeechText-to-speech audio generation | MiniMax | 1 credit (headline rate) | — | Text to audio |
| Chatterbox TTSHigh-quality text-to-speech with Chatterbox | Chatterbox | 3 credits (headline rate) | — | Text to audio |
| Suno Sounds V5variantSuno Sounds V5 generates high-quality sound effects and background music from text prompts with looping, tempo, and key controls | Suno | 2 credits | — | Text to audio |
1 October – 16 December 2025
6 models · 3 providersCartesia, Chatterbox and Inworld shipped 6 audio models between 1 October 2025 and 16 December 2025. Of the 6, 2 have a spec page, 2 are folded into a parent model's page as tier or mode variants and 2 have no page of their own.
- First voice clone model in the catalog — Cartesia Sonic 3.
- New names on the roster: Cartesia, Inworld.
| Model | Built by | Credits per job | Max output | Capabilities |
|---|---|---|---|---|
| Cartesia Sonic 3variantLatest and most capable Cartesia TTS model. Multilingual, expressive, and supports emotion control, speed tuning, and voice… | Cartesia | 4 credits (headline rate) | — | Text to audio, Voice clone |
| Cartesia Voice CloneClone any voice using Cartesia AI. Upload a short audio sample to instantly create a personalized voice for text-to-speech… | Cartesia | 8 credits (headline rate) | — | Voice clone |
| Inworld TTSHigh-quality text-to-speech powered by Inworld AI. Supports a curated library of expressive voices with multilingual capability. | Inworld | 3 credits (headline rate) | — | Text to audio |
| Inworld Voice ClonevariantClone any voice using Inworld AI voice cloning. Upload audio samples to create a personalized voice for TTS generation. | Inworld | 5 credits (headline rate) | — | Text to audio, Voice clone |
| Chatterbox STS TurboChatterbox speech-to-speech turbo voice conversion model | Chatterbox | 9 credits (headline rate) | — | Audio to audio |
| Chatterbox TTS TurboTurbo-speed text-to-speech with Chatterbox | Chatterbox | 8 credits (headline rate) | — | Text to audio |
When 2025 was busy
5 of the twelve months carried a audio release; the busiest window was 1 October – 16 December 2025, with 6.
- April 2025
- 1
- May 2025
- 1
- September 2025
- 1
- October 2025
- 4
- December 2025
- 2
Who shipped audio in 2025
Chatterbox (3), Cartesia (2) and Inworld (2) led on volume. SKUs, not quality — four tiers of one model count four times.
| Provider | Models | Spec pages | First | Latest |
|---|---|---|---|---|
| Chatterbox | 3 | 0 | 1 May 2025 | 16 December 2025 |
| Cartesia | 2 | 2 | 1 October 2025 | 1 October 2025 |
| Inworld | 2 | 1 | 17 October 2025 | 17 October 2025 |
| MiniMax | 1 | 0 | 2 April 2025 | 2 April 2025 |
| Suno | 1 | 1 | 23 September 2025 | 23 September 2025 |
What 2025 moved
Firsts, measured against every audio model released before them. “New name on the roster” is the provider label in the catalog, not the company — one lab can hold several labels.
- 2 April – 23 September 2025
- Longest single clip the catalog had offered — 120s on Chatterbox TTS.
- New names on the roster: MiniMax, Chatterbox, Suno.
- 1 October – 16 December 2025
- First voice clone model in the catalog — Cartesia Sonic 3.
- New names on the roster: Cartesia, Inworld.
Dates are each model’s released_at value, walked in order and grouped until a window held 3 or more. Credits are what one complete generation costs per the model’s own price matrix; where it states none, the headline rate is shown and the model sits out the cheapest-in-window line.
Other release years
Frequently asked questions
How many AI audio models were released in 2025?+
Versely's catalog carries 9 audio models with a 2025 release date, from 5 providers, arriving in 2 launch windows between 2 April 2025 and 16 December 2025.
What was the biggest AI audio launch of 2025?+
1 October – 16 December 2025, with 6 audio models from Cartesia, Chatterbox and Inworld. Of the 6, 2 have a spec page, 2 are folded into a parent model's page as tier or mode variants and 2 have no page of their own.
Which company released the most AI audio models in 2025?+
Chatterbox, with 3 of the 9 audio models dated 2025 — first on 1 May 2025, most recently on 16 December 2025. Cartesia shipped 2, Inworld shipped 2, MiniMax shipped 1.
What changed in AI audio generation in 2025?+
Measured against everything the catalog carried before it: Longest single clip the catalog had offered — 120s on Chatterbox TTS; New names on the roster: MiniMax, Chatterbox, Suno; First voice clone model in the catalog — Cartesia Sonic 3.
Run any 2025 audio model in Versely
The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.