Gemini Flash TTS — Expressive speech, with laughs, sighs and pauses exactly where you put them.
Gemini Flash TTS is Google’s speech model; on Deepnia it runs Gemini 3.8 Flash TTS, generally available since September 22, 2026. You write up to 5,000 characters, pick one of Google’s prebuilt voices and place sounds such as a laugh, a sigh or a pause with English tags, such as <laugh>: the language is recognized from the text and you get a WAV file.
The voice generator opens with Gemini Flash TTS already selected. The exact cost shows before every creation; you sign up when you generate.
<laugh>
<sigh>
<gasp>
<short pause>
Vocal tags Google recommends, written in English between angle brackets.
At a glance
Maker
Google
Model
Gemini 3.8 Flash TTS
Text
Up to 5,000 characters per generation
Language
Automatic detection, 130+ languages according to Google
Voices
Google prebuilt voices, Charon by default
Expression
English vocal tags: laughs, sighs, pauses…
Output
1 WAV file per generation
What Gemini Flash TTS does best
Sounds placed word by word
Google recommends 35 vocal tags for Gemini 3.8 Flash TTS, written in English between angle brackets: <laugh>, <sigh>, <gasp>, <short pause>, <long pause>… The model reads your script as written and plays each sound exactly where its tag sits. On Deepnia, the “Expression” panel adds laughs, sighs, pauses or whispers in one click, and you can type Google’s tags by hand.
More than 130 languages
Write in the language of your choice: Google announces more than 130 languages for Gemini 3.8 Flash TTS, and on Deepnia the language is recognized from your text, with nothing to set.
Voices with a clear style
Choose from Google’s 30 prebuilt voices, each described by a style: Charon (informative), Kore (firm), Puck (upbeat), Sulafat (warm), Achernar (soft). Charon is the default voice.
Built for long scripts
Google presents Gemini 3.8 Flash TTS as its flagship speech model: studio-grade voice fidelity, expressive acting, authentic regional accents and strong stability over long texts. Enough to read a whole script with the same voice.
A voice in three steps
Pick a voice, write the text, generate: the language is recognized automatically and each generation delivers one WAV file (24 kHz, mono), ready for editing.
How to use Gemini Flash TTS on Deepnia
Open the voice generator
The “Try Gemini Flash TTS” button opens Deepnia’s voice generator with the model already selected.
Pick a voice
Browse Google’s prebuilt voices (Charon by default), search by name or style and play their previews.
Write and tag the text
Write up to 5,000 characters, then add sounds from the “Expression” panel or type Google’s tags in English between angle brackets.
Generate
Check the cost shown, then click “Generate voice”: the language is detected automatically.
Gemini Flash TTS on Deepnia: what you can set
Inputs
Text
Text
1 to 5,000 characters per generation
Voices
30 Google prebuilt voices, with search and previews; Charon by default
Language
Automatic detection from the text; 130+ languages according to Google
Tags
“Expression” panel, plus Google’s vocal tags in English between angle brackets
Output
1 WAV file (24 kHz, mono) per generation
Prompt ideas to try
Copy a prompt, open the tool and adapt it to your project.
Bedtime story for a family audio channel
Soft voice
Language: auto
Tags: <yawn>, <long pause>
Once upon a time, there was a little fox who never wanted to sleep. <yawn> But that night, she found a hidden door in the old oak tree. <long pause> Behind it, a whole village glowed with fireflies.
A whole document, text or PDF, read into a single MP3 file.
Frequently asked questions
Which languages does Gemini Flash TTS speak?
Google announces more than 130 languages for Gemini 3.8 Flash TTS, French included. Write your text in the language you want: it is recognized from the text, and the tags are written in English.
How does Gemini recognize the language of my text?
Automatically: just write in the language you want and Gemini recognizes it from the text. Keep one language per script for consistent results.
How many voices does Gemini offer?
30 Google prebuilt voices, each with a style and a preview: Charon (informative), Kore (firm), Puck (upbeat)… Charon is the default voice.
How do I add a laugh, a sigh or a pause?
Add them from the “Expression” panel, or type Google’s tags in English between angle brackets, such as <laugh>, <sigh> or <short pause>, exactly where the sound should happen. Google recommends 35 of them.
Which voice should I pick for an ad, a tutorial or a bedtime story?
Pick by style: Puck (upbeat) for an ad, Charon (informative) for a tutorial, Achernar (soft) for a bedtime story, Sulafat (warm) for a welcome message. Play the previews, then add tags where you want a sound.
Is Gemini Flash TTS good for long narration?
Yes: Google designs Gemini 3.8 Flash TTS for studio-grade voice fidelity and strong stability over long texts. On Deepnia, you generate up to 5,000 characters at once with the same voice.
Which Google model powers Gemini Flash TTS?
Gemini 3.8 Flash TTS, which Google made generally available on September 22, 2026: it reads your texts with Google’s 30 prebuilt voices and understands English vocal tags.
How much text can I convert, and in what format?
Up to 5,000 characters per generation, delivered as one WAV file (24 kHz, mono). For a longer text, generate it in several parts and join them in your editor.
Ready to try Gemini Flash TTS?
The voice generator opens with Gemini Flash TTS already selected. The exact cost shows before every creation; you sign up when you generate.