Qwen · Audio model

    Qwen Audio 3 TTS Flash

    Alibaba's hosted Qwen Audio 3.0 TTS (Flash tier): fast, natural multilingual speech across 10 languages with 44 named voices. The hosted successor to the open-weight Qwen 3 TTS models.

    4 credits per 1,000 characters

    What Qwen Audio 3 TTS Flash is best at

    Text to audio
    Text to SpeechMultilingualFastVoice Library

    Pricing

    Qwen Audio 3 TTS Flash costs 4 credits per 1,000 characters on Versely. It bills per 1,000 characters of input text.

    OptionPrice
    per 1,000 characters4 credits

    Prices are Versely credits. Your plan's credit allowance is on the pricing page.

    Inputs

    Text prompt

    Generates from a written prompt — no media upload required.

    Release

    Qwen Audio 3 TTS Flash was released on July 28, 2026released 1 months ago.

    Other audio models

    Use Qwen Audio 3 TTS Flash inside these tools

    Frequently asked questions

    How much does Qwen Audio 3 TTS Flash cost on Versely?+

    Qwen Audio 3 TTS Flash costs 4 credits per 1,000 characters on Versely.

    What is Qwen Audio 3 TTS Flash best for?+

    Qwen Audio 3 TTS Flash is a Qwen audio model built for text to audio.

    Does Qwen Audio 3 TTS Flash need a voice sample?+

    No. Qwen Audio 3 TTS Flash generates audio from a text prompt alone.

    When was Qwen Audio 3 TTS Flash released?+

    Qwen Audio 3 TTS Flash was released on July 28, 2026 (released 1 months ago).

    Try Qwen Audio 3 TTS Flash inside Versely

    The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync — in your browser or on your phone.