Voice model · Deepnia

Open TTS — Expressive voice-overs with captions and your own voice.

Open TTS is Deepnia’s voice model and the default choice in its voice generator. It turns up to 5,000 characters into a voice-over you direct with 22 expression tags, a 0.70×–1.20× speed range and, if you like, your own voice made from a 10–60 second recording. It is also the only Deepnia voice model that delivers a transcript and SRT captions.

  • 5,000 characters
  • Auto + 23 languages
  • Speed 0.70×–1.20×
  • 22 expression tags
  • Transcript and SRT captions
  • Personal voice
  • Dialogues up to 8 voices

The voice generator opens with Open TTS already selected. The exact cost shows before every creation; you sign up when you generate.

  • [whispering]
  • [emphasis]
  • [laughing]
  • [sighing]
A few of Open TTS’s 22 expression tags on Deepnia.

At a glance

Maker
Deepnia
Text
Up to 5,000 characters per generation
Languages
Auto-detect + 23 languages
Speed
0.70× to 1.20× (1.00× by default)
Expression
22 tags, with optional AI tagging
Voices
Narrator library and personal voices
Output
WAV, with optional transcript and SRT

What Open TTS does best

Tags that direct every sentence

Put a tag right before the words it should color: 8 emotional tones such as [excited], [whispering] or [soft], and 14 audio effects such as [laughing], [sighing] or [clear throat]. The “Add tags with AI” button places them for you, 3 per sentence at most, without changing your words.

Captions ready to publish

Tick “Transcript” or “SRT captions” to download the text and an SRT file timed from the word timings, estimated when they are missing. Each caption holds 8 words or 52 characters at most, a comfortable size on vertical video.

Your own voice

Record or upload 10–60 seconds of clear speech, confirm you have permission to use that voice, then find it in the voice library’s “Mes voix” (My voices) tab, marked as personal.

Pacing under control

Set the speed from 0.70× to 1.20×: a little livelier for a social ad, calmer for a tutorial or an audiobook.

A narrator library

Search for a voice, filter by language (French, English, Spanish) or gender, and listen before you generate. The language menu offers auto-detect plus 23 languages, from Arabic to Japanese.

Dialogues with several voices

In the Story Audio workflow, Open TTS reads dialogues with up to 8 library voices, for a chaptered audiobook or a scene with several characters.

How to use Open TTS on Deepnia

  1. Open the voice generator

    The “Try Open TTS” button opens Deepnia’s voice generator with the model already selected.

  2. Pick a voice

    In the narrator library, filter by language or gender, play the voices, or choose your personal voice under “Mes voix” (My voices).

  3. Write and direct the text

    Write up to 5,000 characters, then add tags from the “Expression” panel or with “Add tags with AI”.

  4. Set it up and generate

    Choose the language and speed, tick the transcript or SRT captions under “Text files”, check the cost shown, then click “Generate voice”.

Open TTS on Deepnia: what you can set

Inputs
Text; authorized personal voice
Text
1 to 5,000 characters per generation
Output
1 WAV file (24 kHz, mono) per generation
Text files
Optional transcript and SRT captions, 8 words or 52 characters per caption at most
Languages
Auto-detect + 23 languages: English, French, Spanish, German, Italian, Brazilian Portuguese, Japanese, Chinese, Korean, Hindi, Arabic, Russian, Turkish, Dutch, Ukrainian, Vietnamese, Indonesian, Thai, Polish, Romanian, Greek, Czech, Finnish
Voices
French, English and Spanish narrators, filterable by gender, plus your personal voices
Speed
0.70× to 1.20× (1.00× by default)
Tags
8 emotional tones and 14 audio effects; AI tagging, 3 per sentence at most
Personal voice
10–60 s of clear speech, with confirmation of your permission
Dialogues
Up to 8 voices in the Story Audio workflow

Prompt ideas to try

Copy a prompt, open the tool and adapt it to your project.

Captioned TikTok voice-over for an online store

  • English
  • Upbeat female voice
  • Speed 1.05×
  • SRT captions
  • Tags: [excited], [whispering]

[excited] Three tricks for parcels people love to open! Kraft paper, a handwritten note… [whispering] and a small surprise tucked at the bottom.

Open the tool

E-learning lesson with a transcript

  • English
  • Calm male voice
  • Speed 0.95×
  • Transcript
  • Tag: [emphasis]

In this lesson, we will build your first monthly budget together. [emphasis] Step one: list every fixed expense, from rent to your phone plan.

Open the tool

Podcast intro

  • English
  • Warm female voice
  • Speed 1.00×
  • Tag: [laughing]

Welcome back to Small Shop, Big Ideas, the show for people who run their own business. [laughing] Yes, we finally fixed the microphone. Today: turning first-time buyers into regulars.

Open the tool

Audiobook chapter

  • English
  • Deep narrator voice
  • Speed 0.90×
  • Tags: [soft], [sighing]

[soft] The lighthouse keeper’s house had been empty for ten years. [sighing] And yet, every night, a light came on in the top-floor window.

Open the tool

Fitness app onboarding in Spanish

  • Español
  • Energetic voice
  • Speed 1.05×

Bienvenido a tu nuevo plan de entrenamiento. Elige tu objetivo y te preparamos tu primera semana de ejercicios en menos de un minuto.

Open the tool

In-store closing announcement in French

  • Français
  • Clear, calm voice
  • Speed 0.95×

Chers clients, notre magasin fermera ses portes dans quinze minutes. Merci de vous diriger vers les caisses. Nous vous souhaitons une excellente soirée.

Open the tool

Tips for better results

  1. Put each tag right before the words it should color, and keep to 3 per sentence at most, as Deepnia’s AI tagging does.
  2. Proofread your text before “Add tags with AI”: the AI only inserts tags from the palette and never changes your words.
  3. Match the speed to the use: around 1.05× for a short social video, around 0.90× for a tutorial or an audience new to the topic.
  4. For a script in French, English or Spanish, pick a narrator from that language’s tab in the voice library.
  5. Tick “SRT captions” under “Text files” to show the SRT download button under the generated voice.
  6. For your personal voice, record 10–60 seconds in a quiet room, with no music or echo, in the tone of your future videos.

Good to know

  • Up to 5,000 characters per generation: split a longer text into several parts and generate them one after another.
  • For a dialogue with several voices, use the Story Audio workflow: up to 8 library voices in one scene.
  • Create your personal voice from your own voice, or from someone who has given you their consent.
  • Pick your tags from the “Expression” panel: it gathers the palette’s 8 emotional tones and 14 audio effects.

Open TTS or another model?

Deepnia’s other voice and music models, to pick the right one for your project.

ModelInputsSettingsBest for
Open TTSThis modelText, Authorized personal voiceLanguage, Speed, Expression tagsExpressive voice-overs with captions and your own voice.
ElevenLabsTextLanguage, Stability, Expression tagsA voice from the ElevenLabs catalogue, with three stability levels.
Gemini Flash TTSTextExpression tagsA Google voice steered by tags, with the language detected automatically.
MiniMax Speech 2.6 HDText, Authorized personal voiceLanguage, SpeedA cloned voice, with adjustable speed and 23 languages + Auto.
Deepnia Long-formText, PDF, Authorized personal voice—A whole document, text or PDF, read into a single MP3 file.

Frequently asked questions

What is Open TTS?

It is Deepnia’s voice model and the default in the voice generator. It turns up to 5,000 characters into a WAV voice-over, with expression tags, adjustable speed, personal voices and an optional transcript and SRT captions.

Which languages does Open TTS speak?

The language menu offers auto-detect plus 23 languages, including English, French, Spanish, Arabic and Hindi, and the voice library has French, English and Spanish narrators.

How do I get SRT captions for my voice-over?

Open “Text files” and tick “SRT captions” (and “Transcript” if needed): the download buttons appear under the generated voice. The SRT follows the word timings, with 8 words or 52 characters per caption at most.

How do I add emotions or laughter?

Click a tag in the “Expression” panel (8 emotional tones, 14 audio effects), such as [excited] or [laughing], or let “Add tags with AI” place them: 3 per sentence at most, without rewriting your text.

Can I use my own voice?

Yes. Create a personal voice: record or upload 10–60 seconds of clear speech and tick the box confirming you have permission to use that voice. It then appears under “Mes voix” (My voices). Only clone a voice you have the right to use.

Can I change the speaking speed?

Yes, from 0.70× to 1.20×, with 1.00× by default.

How much text can I convert at once?

Up to 5,000 characters per generation, delivered as one WAV file. Split longer texts into several parts.

Can several voices speak in the same audio?

Yes, with the Story Audio workflow: a dialogue can bring together up to 8 library voices. In the voice generator, pick one voice per generation.

Ready to try Open TTS?

The voice generator opens with Open TTS already selected. The exact cost shows before every creation; you sign up when you generate.

Try Open TTS

Information checked on September 27, 2026.