Voice model · Google

Gemini Flash TTS — Expressive speech, with laughs, sighs and pauses exactly where you put them.

Gemini Flash TTS is Google’s speech model; on Deepnia it runs Gemini 3.8 Flash TTS, generally available since September 22, 2026. You write up to 5,000 characters, pick one of Google’s prebuilt voices and place sounds such as a laugh, a sigh or a pause with English tags, such as <laugh>: the language is recognized from the text and you get a WAV file.

  • 5,000 characters
  • Automatic language detection
  • English vocal tags
  • 30 Google prebuilt voices
  • 130+ languages
  • WAV file

The voice generator opens with Gemini Flash TTS already selected. The exact cost shows before every creation; you sign up when you generate.

  • <laugh>
  • <sigh>
  • <gasp>
  • <short pause>
Vocal tags Google recommends, written in English between angle brackets.

At a glance

Maker
Google
Model
Gemini 3.8 Flash TTS
Text
Up to 5,000 characters per generation
Language
Automatic detection, 130+ languages according to Google
Voices
Google prebuilt voices, Charon by default
Expression
English vocal tags: laughs, sighs, pauses…
Output
1 WAV file per generation

What Gemini Flash TTS does best

Sounds placed word by word

Google recommends 35 vocal tags for Gemini 3.8 Flash TTS, written in English between angle brackets: <laugh>, <sigh>, <gasp>, <short pause>, <long pause>… The model reads your script as written and plays each sound exactly where its tag sits. On Deepnia, the “Expression” panel adds laughs, sighs, pauses or whispers in one click, and you can type Google’s tags by hand.

More than 130 languages

Write in the language of your choice: Google announces more than 130 languages for Gemini 3.8 Flash TTS, and on Deepnia the language is recognized from your text, with nothing to set.

Voices with a clear style

Choose from Google’s 30 prebuilt voices, each described by a style: Charon (informative), Kore (firm), Puck (upbeat), Sulafat (warm), Achernar (soft). Charon is the default voice.

Built for long scripts

Google presents Gemini 3.8 Flash TTS as its flagship speech model: studio-grade voice fidelity, expressive acting, authentic regional accents and strong stability over long texts. Enough to read a whole script with the same voice.

A voice in three steps

Pick a voice, write the text, generate: the language is recognized automatically and each generation delivers one WAV file (24 kHz, mono), ready for editing.

How to use Gemini Flash TTS on Deepnia

  1. Open the voice generator

    The “Try Gemini Flash TTS” button opens Deepnia’s voice generator with the model already selected.

  2. Pick a voice

    Browse Google’s prebuilt voices (Charon by default), search by name or style and play their previews.

  3. Write and tag the text

    Write up to 5,000 characters, then add sounds from the “Expression” panel or type Google’s tags in English between angle brackets.

  4. Generate

    Check the cost shown, then click “Generate voice”: the language is detected automatically.

Gemini Flash TTS on Deepnia: what you can set

Inputs
Text
Text
1 to 5,000 characters per generation
Voices
30 Google prebuilt voices, with search and previews; Charon by default
Language
Automatic detection from the text; 130+ languages according to Google
Tags
“Expression” panel, plus Google’s vocal tags in English between angle brackets
Output
1 WAV file (24 kHz, mono) per generation

Prompt ideas to try

Copy a prompt, open the tool and adapt it to your project.

Bedtime story for a family audio channel

  • Soft voice
  • Language: auto
  • Tags: <yawn>, <long pause>

Once upon a time, there was a little fox who never wanted to sleep. <yawn> But that night, she found a hidden door in the old oak tree. <long pause> Behind it, a whole village glowed with fireflies.

Open the tool

Radio spot for a sports store

  • Upbeat voice
  • Language: auto
  • Tags: <cheer>, <short pause>

<cheer> This weekend, it’s running week at your local sports store! <short pause> Shoes, watches and gear are waiting for you until Sunday night.

Open the tool

Guided meditation

  • Calm voice
  • Language: auto
  • Tags: <breath>, <long pause>

Breathe in deeply… <breath> then let your shoulders drop. <long pause> Let your breath find its own rhythm.

Open the tool

Smart-device tutorial

  • Voice: Charon
  • Language: auto
  • Tag: <short pause>

Press the green button, then wait three seconds before speaking. <short pause> The blue light confirms that recording has started.

Open the tool

Podcast intro in French

  • Warm voice
  • Language: auto (French)
  • Tag: <laugh>

Bienvenue dans Les Bricoleurs du week-end, le podcast des projets perso devenus sérieux. <laugh> Promis, cet épisode est plus court que le précédent.

Open the tool

Answer in a quiz app

  • Upbeat voice
  • Language: auto
  • Tags: <sigh>, <short pause>

So close! <sigh> The right answer was seven continents. <short pause> Next question: what is the longest river in the world?

Open the tool

Tips for better results

  1. Write tags in English, even inside a script in another language: that is Google’s guidance.
  2. Put each tag exactly where the sound should happen: Gemini 3.8 reads your script as written.
  3. Set the pace with <short pause> or <long pause>, placed between two sentences.
  4. Write each script in a single language: Gemini recognizes it automatically.
  5. Pick the voice for its style first (informative, firm, upbeat, warm…), then add tags.

Good to know

  • For a text longer than 5,000 characters, generate it in several parts with the same voice, then join the files in your editor.
  • In the “Expression” panel, favor the sounds Google documents: laughs, sighs, pauses and whispers.
  • For a series of tutorials or episodes, keep the same voice and the same style of tags from file to file: the tone stays consistent.
  • For a story with several characters, Open TTS’s Story Audio workflow brings together up to 8 voices.

Gemini Flash TTS or another model?

Deepnia’s other voice and music models, to pick the right one for your project.

ModelInputsSettingsBest for
Gemini Flash TTSThis modelTextExpression tagsExpressive speech, with laughs, sighs and pauses exactly where you put them.
Open TTSText, Authorized personal voiceLanguage, Speed, Expression tagsAdjustable speed, a personal voice, 22 tags and SRT captions.
ElevenLabsTextLanguage, Stability, Expression tagsA 10-language menu and three stability levels.
MiniMax Speech 2.6 HDText, Authorized personal voiceLanguage, SpeedA cloned voice, with adjustable speed and 23 languages + Auto.
Deepnia Long-formText, PDF, Authorized personal voice—A whole document, text or PDF, read into a single MP3 file.

Frequently asked questions

Which languages does Gemini Flash TTS speak?

Google announces more than 130 languages for Gemini 3.8 Flash TTS, French included. Write your text in the language you want: it is recognized from the text, and the tags are written in English.

How does Gemini recognize the language of my text?

Automatically: just write in the language you want and Gemini recognizes it from the text. Keep one language per script for consistent results.

How many voices does Gemini offer?

30 Google prebuilt voices, each with a style and a preview: Charon (informative), Kore (firm), Puck (upbeat)… Charon is the default voice.

How do I add a laugh, a sigh or a pause?

Add them from the “Expression” panel, or type Google’s tags in English between angle brackets, such as <laugh>, <sigh> or <short pause>, exactly where the sound should happen. Google recommends 35 of them.

Which voice should I pick for an ad, a tutorial or a bedtime story?

Pick by style: Puck (upbeat) for an ad, Charon (informative) for a tutorial, Achernar (soft) for a bedtime story, Sulafat (warm) for a welcome message. Play the previews, then add tags where you want a sound.

Is Gemini Flash TTS good for long narration?

Yes: Google designs Gemini 3.8 Flash TTS for studio-grade voice fidelity and strong stability over long texts. On Deepnia, you generate up to 5,000 characters at once with the same voice.

Which Google model powers Gemini Flash TTS?

Gemini 3.8 Flash TTS, which Google made generally available on September 22, 2026: it reads your texts with Google’s 30 prebuilt voices and understands English vocal tags.

How much text can I convert, and in what format?

Up to 5,000 characters per generation, delivered as one WAV file (24 kHz, mono). For a longer text, generate it in several parts and join them in your editor.

Ready to try Gemini Flash TTS?

The voice generator opens with Gemini Flash TTS already selected. The exact cost shows before every creation; you sign up when you generate.

Try Gemini Flash TTS

Information checked on September 28, 2026.