Voice model · MiniMax

MiniMax Speech 2.6 HD — HD voice-overs, in your own voice if you like.

MiniMax Speech 2.6 HD is MiniMax’s high-definition text-to-speech model, released at the end of October 2025. On Deepnia it reads up to 5,000 characters per generation in 23 languages or with auto-detect, at a speed you set from 0.70× to 1.20×. Its strength here: speaking in your own voice, cloned with your consent from a 10–60 second recording.

  • 5,000 characters
  • 23 languages + auto
  • Speed 0.70×–1.20×
  • Cloned voice (10–60 s)
  • HD voice-over
  • WAV file

The voice generator opens with MiniMax Speech 2.6 HD already selected. The exact cost shows before every creation; you sign up when you generate.

  • Auto
  • FR
  • EN
  • ES
  • 0.70×–1.20×
Illustration of the settings Deepnia offers: language (or auto-detect) and speed.

At a glance

Maker
MiniMax
Text
Up to 5,000 characters per generation
Languages
23 + auto-detect
Speed
0.70× to 1.20× (1.00× by default)
Voices
MiniMax voices or your cloned voice
Output
1 WAV file (24 kHz, mono)

What MiniMax Speech 2.6 HD does best

Your voice, cloned with your consent

In the voice clone tool, record or upload 10 to 60 seconds of clear speech and confirm you have permission to use the voice. It then appears at the top of the MiniMax voice list, with your personal voices. MiniMax says its Fluent LoRA technique keeps a speaker’s timbre even when the recording is accented or hesitant.

Everyday formats, read as written

MiniMax says Speech 2.6 reads web addresses, emails, phone numbers, dates and amounts without you spelling them out. Handy for a phone greeting, an appointment reminder or a delivery message.

23 languages plus auto-detect

Pick the language of your text from 23, including English, French, Spanish, Arabic, Hindi and Brazilian Portuguese, or let auto-detect decide.

Faithful HD sound

MiniMax designed the HD version of Speech 2.6 for audio quality and similarity to the original voice: an asset for a polished voice-over and for your cloned voice.

Pacing you control

Set the speed from 0.70× to 1.20× in 0.01 steps. Go slower for a number or opening hours people need to note, faster for an upbeat announcement.

How to use MiniMax Speech 2.6 HD on Deepnia

  1. Create your voice (optional)

    In the voice clone tool, record or upload 10–60 seconds of speech, tick the permission box, then start the clone.

  2. Open the voice generator

    The “Try MiniMax Speech 2.6 HD” button opens Deepnia’s voice tool with the model already selected.

  3. Pick a voice and paste your text

    Choose a MiniMax voice or one of your own, then paste up to 5,000 characters into “Text to speak”.

  4. Set it up and generate

    Choose the language and speed, check the cost shown, then click “Generate voice”.

MiniMax Speech 2.6 HD on Deepnia: what you can set

Inputs
Text; a personal voice made in the voice clone tool
Text
1 to 5,000 characters per generation
Languages
23 languages plus auto-detect, including English, French, Spanish, German, Italian, Brazilian Portuguese, Arabic, Hindi, Japanese, Chinese and Korean
Speed
0.70× to 1.20× in 0.01 steps, 1.00× by default
Voices
A selection of MiniMax voices with search; your cloned voices at the top of the list
Cloned voice
A 10–60 s sample, recorded or uploaded; language (auto or 11 languages), noise reduction and volume normalization; permission required
Mode
HD voice-over
Output
1 WAV file (24 kHz, mono) per generation

Prompt ideas to try

Copy a prompt, open the tool and adapt it to your project.

Phone greeting for a dental practice

  • Calm voice
  • English
  • Speed 0.95×

Thank you for calling Linden Dental. We are open Monday to Friday, from 8:30 a.m. to 6:30 p.m. To book online, visit example.com/book. For emergencies, call (202) 555-0147. Please hold, and we will be with you shortly.

Open the tool

Shipping update for an online store

  • Upbeat voice
  • English
  • Speed 1.05×

Hi Jordan! Your order #48213 left our warehouse on November 3 and should arrive between November 6 and 8. Track it at example.com/track or email us at help@example.com.

Open the tool

Appointment reminder for a hair salon

  • Warm voice
  • English
  • Speed 1.00×

Hello, this is Lumière Salon with a reminder of your appointment on March 12 at 2:45 p.m. with Ines. To reschedule, call (202) 555-0147 at least 24 hours ahead. See you soon!

Open the tool

Podcast intro in your own voice

  • Cloned voice
  • English
  • Speed 1.00×

Welcome to Road Notes, episode 42. I’m Leah, and today we’re exploring night markets around the world. Show notes are at example.com/podcast. Let’s get started!

Open the tool

Closing announcement for a supermarket

  • Calm voice
  • English
  • Speed 0.90×

Attention, shoppers: the store will close in 15 minutes, at 9 p.m. Please make your way to the checkouts. We open again tomorrow at 8:30 a.m. Thank you for shopping with us.

Open the tool

French welcome message for a hotel in Kigali

  • Warm voice
  • French
  • Speed 1.00×

Bienvenue au Hillside Hotel de Kigali. Le petit-déjeuner est servi en terrasse de 6 h 30 à 10 h. Pour le service en chambre, composez le 9. Une question avant votre arrivée ? Écrivez-nous à stay@example.com.

Open the tool

Tips for better results

  1. Keep numbers, dates and web addresses in their usual form, such as 555-0147, March 12 or example.com: MiniMax says it reads them without preparation. Listen to a first take before you publish.
  2. For your cloned voice, record 10–60 seconds in a quiet room with no background music, and keep noise reduction and volume normalization switched on.
  3. Pick the language of your text when you know it: MiniMax uses it to recognize the language better. Auto-detect is still available.
  4. Use punctuation to shape the rhythm: commas and full stops mark pauses.
  5. Slow down a little (0.90×–0.95×) for a number or time people must note; speed up slightly for an upbeat announcement.
  6. Beyond 5,000 characters, split the script into several generations at paragraph breaks, keeping the same voice and speed.

Good to know

  • Clone your own voice or that of someone who has given you their consent: confirming you have permission to use the voice is required to start the clone.
  • Once created, your voice appears at the top of the MiniMax voice list, with your personal voices: pick it for every new message.
  • For a phone greeting or a series of announcements, keep the same voice and speed from one message to the next: your brand keeps a recognizable voice.

MiniMax Speech 2.6 HD or another model?

Deepnia’s other voice and music models, to pick the right one for your project.

ModelInputsSettingsBest for
MiniMax Speech 2.6 HDThis modelText, Authorized personal voiceLanguage, SpeedHD voice-overs, in your own voice if you like.
Open TTSText, Authorized personal voiceLanguage, Speed, Expression tagsExpression tags, plus a transcript or SRT captions with the audio.
ElevenLabsTextLanguage, Stability, Expression tagsElevenLabs catalogue voices and three stability levels, in 10 languages or auto.
Gemini Flash TTSTextExpression tagsExpression tags with automatic language detection.
Deepnia Long-formText, PDF, Authorized personal voice—A whole document (PDF, Word, PowerPoint) read into a single MP3 file.

Frequently asked questions

Does MiniMax Speech 2.6 HD speak French?

Yes. French is in Deepnia’s language menu, and MiniMax documents French among the model’s languages. Listen to a short test to check how numbers and acronyms are read.

Can I use my own voice?

Yes. In the voice clone tool, record or upload 10–60 seconds of clear speech, confirm you have permission to use the voice, then pick it at the top of the MiniMax voice list.

How many languages are available?

23 on Deepnia, plus auto-detect, including English, French, Spanish, German, Arabic, Hindi, Japanese and Brazilian Portuguese.

What is the maximum text length?

5,000 characters per generation on Deepnia. For a longer script, split it into several generations.

Can I change the speed?

Yes, from 0.70× to 1.20× in 0.01 steps, 1.00× by default.

How do I get a natural read?

Use punctuation to shape the rhythm (commas and full stops mark pauses), keep numbers, dates and web addresses in their usual form, pick the language of your text and adjust the speed: a little slower for information people must note, a little faster for an upbeat announcement.

Which voice should I pick for an ad?

An upbeat voice set around 1.05× gives an announcement some drive; your own cloned voice embodies your brand. Listen to a short test, then fine-tune the speed between 0.70× and 1.20×.

What format is the audio?

One WAV file (24 kHz, mono) per generation, ready to download.

Ready to try MiniMax Speech 2.6 HD?

The voice generator opens with MiniMax Speech 2.6 HD already selected. The exact cost shows before every creation; you sign up when you generate.

Try MiniMax Speech 2.6 HD

Information checked on September 27, 2026.