Gemini 3.8 Flash TTS 🧠 Model Paid
Google's prompt-directed TTS with 2,000+ voices, 100+ languages and consented cloning
- Maker
- Google DeepMind
- Latest release
- 2026-09
- Weights
- Closed
- Licence
- –
- Apps offering it
- 1
Where you can use Gemini 3.8 Flash TTS
Apps in this directory that let you generate with Gemini 3.8 Flash TTS, from their own documentation. You can also call it directly through the maker's API.
Good for
About Gemini 3.8 Flash TTS
What it is
Gemini TTS is Google's speech generation line in the Gemini API. On 23 September 2026 Google launched two new models. Gemini 3.8 Flash TTS is built for top quality and directed performances. Gemini 3.8 Flash-Lite TTS is built for high-volume dubbing and agents. Both take natural-language direction for delivery and replace the earlier 3.1 Flash and 2.5 Pro TTS previews.
Key features
- Over 100 languages, with regional varieties such as Quebec French and Scots English
- 30 core studio voices plus a library of 2,000+ ready-made voices
- Custom voices designed from a text description
- Voice replication from a 30-second sample, with a verified consent recording
- Two-speaker dialogue and line-by-line delivery direction
Where you can use it
The Gemini API and Google AI Studio. Enterprise API access is coming soon. For end users, Flash TTS is rolling out in Gemini Notebook and Flash-Lite TTS in Google Vids. Neither is a separate listing in our directory yet.
Pricing and rights
API pricing per million tokens, valid until 31 December 2026: Flash TTS costs $0.50 input and $9 audio output. Flash-Lite TTS costs $0.50 input and $6 output. Both prices double from 1 January 2027. All audio carries an inaudible SynthID watermark. Voice replication needs a spoken consent recording from the voice owner, and it is not offered in Illinois, Texas, the EEA, the UK, Switzerland or India.
Who it is for
Developers and media teams producing narration, dubbing or multi-voice content at scale, especially in many languages.
Verdict
Gemini 3.8 Flash TTS combines a huge voice library, broad language coverage and strong consent safeguards. It is brand new, prices double in January 2027, and cloning is blocked in several major markets.
Pros
- 2,000+ voices and 100+ languages
- Natural-language delivery direction
- SynthID watermark on all audio
- Consent-verified voice cloning
Cons
- Launched 23 Sept 2026, still very new
- Prices double from January 2027
- Cloning unavailable in the EU, UK, India and some US states
Similar AI models
All speech & voice models →Eleven v3 🧠 ModelFreemium
ElevenLabs' expressive text-to-speech model with audio tags and 70+ languages
Whisper large-v3-turbo 🧠 ModelOpen source
OpenAI's open-source speech recognition for transcripts and subtitles in 99 languages
Chatterbox Multilingual V3 🧠 ModelOpen source
Resemble AI's MIT-licensed TTS with zero-shot voice cloning and built-in watermarking
Sonic-3.6 🧠 ModelFreemium
Cartesia's low-latency TTS for voice agents and narration, 44 languages
Kokoro-82M v1.0 🧠 ModelOpen source
Tiny Apache-licensed TTS model with 54 voices in 8 languages that runs anywhere
GPT-4o mini TTS 🧠 ModelPaid
OpenAI's steerable text-to-speech API with 13 voices and prompt-based style control
More from Google DeepMind
Nano Banana 2 🧠 ModelFreemium
Google's Gemini image models (Nano Banana 2, Pro and 2 Lite) for generation and editing up to 4K
Veo 3.1 🧠 ModelPaid
Google DeepMind's text- and image-to-video model with native sound, up to 4K
Gemini 3.1 Pro 🧠 ModelFreemium
Google's Gemini models (3.1 Pro, 3.8 Flash) in the Gemini app, NotebookLM and the Gemini API
Lyria 3.5 🧠 ModelPaid
Google DeepMind's music model for full songs with vocals, lyrics and SynthID