A whole document, one MP3
Upload a PDF, a Word document (DOCX) or a PowerPoint deck (PPTX) up to 50 MB, or paste a long text: you get a single MP3 file (44.1 kHz, 128 kbps) to play or download.
Voice model · Deepnia
Coming soon to everyone
Deepnia Long-form is Deepnia’s own long-form narration tool: it turns a PDF, a Word document, a PowerPoint deck (up to 50 MB) or a long pasted text into a single MP3 file. The text is extracted in reading order, never summarized or rewritten; you review the full transcript and choose the voice before any audio is made. Deepnia Long-form is coming soon: the tool is already visible as a preview.
Deepnia Long-form is coming soon: the tool already opens as a preview. The exact cost, recalculated from the approved transcript, will show before every creation.
Upload a PDF, a Word document (DOCX) or a PowerPoint deck (PPTX) up to 50 MB, or paste a long text: you get a single MP3 file (44.1 kHz, 128 kbps) to play or download.
The text is extracted in reading order, with no summary, addition or rewording. Only repeated headers and footers, page numbers, images and words split at line ends are removed; table rows are read cell by cell.
You read and correct the full transcript, with search. Uncertain passages from text recognition are flagged with their page and a confidence level, and each one can be marked as checked.
The tool shows every stage: document analysis, text extraction, transcript to approve, audio creation, MP3 assembly. You can close the page: processing continues in the background and a notification tells you when the audio is ready.
Pick a voice from a library of English, French and Spanish narrators, filter by gender and listen to samples first. The tool also shows a personal voices tab for a voice created with your permission.
Your upload stays a private file. 24 hours after the job ends, the source document and intermediate files are deleted; only the MP3 is kept.
Drop a PDF, DOCX or PPTX (50 MB at most) or paste your text: the AI extracts the text in reading order.
Fix the text, check the flagged passages, then save: every correction recalculates the cost.
Listen to samples, filter by language or gender, then approve the exact cost, which is locked at approval.
The AI creates a single MP3 to download or open in your Library. You can also cancel the job along the way.
Copy a prompt, open the tool and adapt it to your project.
A 40-page annual report with a few tables of figures, for members to listen to before the general meeting.
30 pages of biology course notes to replay on the bus while revising for exams.
A 25-slide training deck on welcoming customers, for new hires to listen to before their first week.
The first chapter of a novel, pasted as plain text, to hear the pacing of the story before preparing an audiobook.
An 8-page monthly newsletter in Spanish, for club members who would rather listen than read.
A user guide for a home appliance, in French, offered as audio to customers with low vision or who prefer to listen.
Deepnia’s other voice and music models, to pick the right one for your project.
| Model | Inputs | Settings | Best for |
|---|---|---|---|
| Deepnia Long-formThis model | Text, PDF, Authorized personal voice | — | A whole document, read faithfully, in one MP3. |
| Open TTS | Text, Authorized personal voice | Language, Speed, Expression tags | Text up to 5,000 characters with expression tags, speed and SRT captions. |
| MiniMax Speech 2.6 HD | Text, Authorized personal voice | Language, Speed | A voice-over up to 5,000 characters in your own cloned voice. |
| ElevenLabs | Text | Language, Stability, Expression tags | A short voice-over with expression tags, in 10 languages or auto. |
| Gemini Flash TTS | Text | Expression tags | A short voice-over with expression tags and automatic language detection. |
PDFs, Word documents (DOCX) and PowerPoint decks (PPTX) up to 50 MB, in the document-to-audio workflow. In the voice tool, Deepnia Long-form takes a PDF or pasted text.
50 MB, and 500 pages for a PDF. The approved text can reach 1,000,000 bytes.
Yes. The text is read as written, in reading order: only repeated headers and footers, page numbers, images and words split at line ends are removed, with no summary, addition or rewording.
Yes. The full transcript is editable and searchable, and uncertain passages are flagged with their page and confidence level. Every correction recalculates the cost before approval.
A single MP3 file (44.1 kHz, 128 kbps), ready to download or open in your Library.
Yes. Processing continues in the background and a notification appears when the audio is ready. You can also cancel the job.
The library offers English, French and Spanish narrators: pick a voice in the document’s language.
Deepnia Long-form is coming soon: the tool is already visible as a preview, marked “Bientôt disponible” (coming soon), and will open to everyone shortly. This page describes the workflow as the tool presents it.
Deepnia Long-form is coming soon: the tool already opens as a preview. The exact cost, recalculated from the approved transcript, will show before every creation.