Gemini 3.1 Flash TTS
Gemini 3.1 Flash TTS supports documented multi-speaker dialogue and voice controls, but a reliable maximum output-audio estimate is not yet available. Its details remain available while generation stays disabled.
Online use is currently unavailable
This page is for reviewing the model's capabilities, use cases, and public references. The model does not appear in the home workbench, and you cannot submit generation tasks with it.
At a glance
What is this model like?
This model is neither Coming Soon nor discontinued. The current restriction comes from this site's preauthorization safety boundary: the speaker and dialogue arrays have no total item limit, and there is no provable worst-case conversion from input text to output audio tokens. The model is therefore excluded from selectors, and every request is rejected before any credits are deducted.
Key facts
Quickly assess whether this model fits your use case.
Also known as
The same model may appear under different names across documentation and community discussions; this list keeps them easy to search and compare.
Common questions
These practical questions focus on the task, input conditions, and result requirements you should confirm before choosing.
Selection guide
Use these decision points when choosing a model.
Use ElevenLabs Dialogue V3 first, and test speaker separation with a short script.
Compare ElevenLabs Turbo 2.5 with Multilingual V2 first.
Save three non-sensitive scripts: narration, a two-person interview, and a short character scene.
Popular comparisons
Compare common alternatives on the same task to clarify differences in inputs, controls, and cost.
Model comparison
Compare the current model with alternatives at a glance.
Practical usage insights
Practical guidance based on public sources, current on-site limits, and representative tasks.
Low latency does not establish a cost ceiling
Multi-character input expands worst-case usage
Capabilities
Multi-Speaker Configuration
Ordered Dialogue
Voice Direction
Use cases
Podcast Dialogue Planning
Character Scene Preparation
Future Side-by-Side Testing
Prompt tips
Keep Speaker IDs Consistent
Validate with a Short Script First
Do Not Estimate Audio Length
Why choose it
What to know first
Currently unavailable
This page is for reviewing the model's capabilities, use cases, and public references. The model does not appear in the home workbench, and you cannot submit generation tasks with it. A Gemini TTS candidate designed for faster speech generation and controlled multi-character dialogue.
FAQ
Related models
Compare these similar candidates before deciding.
ElevenLabs Dialogue V3
ElevenLabs Dialogue V3 provides line-by-line dialogue generation: enter each turn, choose a voice, and pay by total dialogue length. It is not a real-time voice agent, and ElevenLabs notes that users may need multiple generations to find a usable result. Start with a sample to check speaker changes, audio tags, language code, output format, and long-script completeness.
ElevenLabs Turbo 2.5
ElevenLabs Turbo 2.5 is a retained single-speaker TTS workflow on Kyeo AI. It exposes a voice choice plus stability, similarity boost, style, speed, timestamps, surrounding text, and language code. The workflow remains selectable, but ElevenLabs now recommends Flash v2.5 instead. The Turbo name does not make low latency, language support, timestamp accuracy, voice authorization, or output quality a Kyeo guarantee.
ElevenLabs Multilingual V2
ElevenLabs Multilingual V2 is Kyeo AI's current single-speaker, high-naturalness voiceover page. It shares a similar form with Turbo 2.5, but the task boundary differs: Turbo is a fast voiceover baseline, while Multilingual V2 is aimed at long-form narration, cross-language content, and sustained brand voice. Kyeo still uses a fixed voice list and a 5,000-character entry, so the practical decision is whether the listening experience justifies twice Turbo's per-character cost.
Sources
Content is based on model documentation, feature references, and the settings available on Kyeo AI.
The speaker, dialogue, voice-control, and token-rate contract for Gemini 3.1 Flash TTS has been verified. However, the maximum number of output audio tokens cannot be proven before a request is created, so generation is not currently available on this site.