Google · Audio model

    Gemini 3.1 Flash TTS

    Google Gemini 3.1 Flash TTS — expressive text-to-speech with 30 voices, natural-language style control, inline audio tags ([sigh], [laughing], [whispering]), multilingual synthesis, and multi-speaker dialogue.

    4 credits#3 overall

    What Gemini 3.1 Flash TTS is best at

    Text to audio
    Text to speechExpressiveAudio tagsMultilingualMulti speaker30 voices

    Pricing

    Gemini 3.1 Flash TTS costs 4 credits on Versely. It bills per 1,000 characters of input text.

    OptionPrice
    per 1,000 characters12 credits

    Prices are Versely credits. Your plan's credit allowance is on the pricing page.

    Inputs

    Text prompt

    Generates from a written prompt — no media upload required.

    Rankings

    Gemini 3.1 Flash TTS ranks #3 overall with an Elo score of 1210 on Versely's live model rankings.

    CategoryRankEloMeasured
    Text to audio#31210

    Release

    Gemini 3.1 Flash TTS was released on April 15, 2026released 4 months ago.

    Compare Gemini 3.1 Flash TTS

    Side-by-side pricing, resolution and rankings against models buyers weigh it against.

    Other audio models

    Use Gemini 3.1 Flash TTS inside these tools

    Frequently asked questions

    How much does Gemini 3.1 Flash TTS cost on Versely?+

    Gemini 3.1 Flash TTS costs 4 credits on Versely.

    What is Gemini 3.1 Flash TTS best for?+

    Gemini 3.1 Flash TTS is a Google audio model built for text to audio.

    How does Gemini 3.1 Flash TTS rank against other audio models?+

    Gemini 3.1 Flash TTS ranks #3 overall with an Elo score of 1210 on Versely's live model rankings.

    Does Gemini 3.1 Flash TTS need a voice sample?+

    No. Gemini 3.1 Flash TTS generates audio from a text prompt alone.

    When was Gemini 3.1 Flash TTS released?+

    Gemini 3.1 Flash TTS was released on April 15, 2026 (released 4 months ago).

    Try Gemini 3.1 Flash TTS inside Versely

    The all-in-one AI studio for creators. 60+ models for video, image, voice, music and lipsync in a single app.