Frequently Asked Questions (FAQs)

Browse common questions about Unmixr. Can't find what you need? Reach us at support@unmixr.com.

Voiceover

What is the most realistic AI text to speech tool?

Unmixr generates some of the most realistic AI voiceovers available, with a voice for every use case across 100+ languages and 150+ regional accents, plus per-line control over emotion, pacing, pitch, and pronunciation. Unlike a basic TTS generator, the same project also produces captions, video narration, and multilingual dubs — so the voiceover arrives publish-ready.

Can I use Unmixr voiceovers commercially?

Yes. On any purchased plan, every audio file you generate is 100% cleared for commercial use — courses, ads, videos, podcasts, and client work — with no separate licence to buy and no royalty per export. Free-trial audio is meant for evaluation.

How many voices and languages does Unmixr support?

Unmixr covers 100+ languages and 150+ regional accents with a voice for every use case, including multilingual voices that automatically detect and switch languages within one script. Voice cloning is supported in 80+ languages.

Can I control emotion, pacing, and pronunciation?

Yes. Many voices support pre-trained emotions (cheerful, sad, excited, whispering, and more) with adjustable intensity, and every voice can be directed with speaking rate, pitch, volume, pauses, emphasis, and breathing. Pronunciation rules for names and acronyms are saved to a library and applied project-wide.

What formats can I export — MP3, SRT, or video?

Yes. Studio audio downloads as a mastered MP3 at 192 kbps. Dubbing Studio additionally delivers a WAV alongside the MP3, and the developer API can return either. The same project also generates matching subtitles and transcripts, and can continue into PowerPoint-to-Video or Training Video Maker to export a narrated MP4.

How is Unmixr different from a basic TTS generator?

A TTS generator hands you an audio file and stops. Unmixr is an AI voice studio: the script that produced your voiceover also produces captions, a narrated video, and dubbed versions in 100+ languages — in one project, without re-recording. Voice quality is the starting point; the publishing workflow is the difference.

Can I keep the same voice across a whole course or video series?

Yes. Pick one AI voice for the entire curriculum, or clone an approved voice (supported in 80+ languages) so every lesson, module, and update sounds like the same narrator. Saved presets and pronunciation rules keep delivery consistent even months later.

How do credits work in text-to-speech?

Credits are calculated from the number of characters in your script. Standard voices (prebuilt, multilingual, cloned) cost 1 credit per character; Persona voices cost 2 credits per character because they use more advanced AI. Each plan includes a monthly or lifetime credit balance.

Can I use multiple voices in one script?

Yes. Assign different voices to different parts of the same script — ideal for dialogue, interviews, and character-driven storytelling. The Dialogue Studio adds drag-and-drop reordering for multi-speaker scripts.

How do I add pauses and sound effects to a voiceover?

Add pauses by selecting text and choosing the pause option, or by typing a tag like [pause 8s]. Sound effects and background music can be inserted before or after any selected text, and you can upload your own effects.

How do I import a script into Unmixr?

Paste your script directly into the editor, or use the Import wizard to bring in PDF, Word, ePub, and text files — or any public webpage.

Can I generate and edit subtitles for my voiceover?

Yes. Unmixr generates subtitles that match your narration, and you can edit them to correct any word or adjust timing before export.

What is an AI voice generator?

An AI voice generator turns written text into spoken audio using a synthetic voice. In Unmixr, the generated voice remains part of an editable project, so you can change the script, direct the delivery, assign characters, fix pronunciation, add music, and continue into captions, video, or dubbing without starting again.

Are the AI voice examples real generated audio?

Yes. The examples on this page are finished generated audio, not actors describing the product. You can play each result without signing up, read the exact transcript and creative direction, and compare product, training, and storytelling performances.

How do I make an AI voice sound lively instead of monotonous?

Start with a voice suited to the use case, then direct the performance. In Unmixr you can describe the mood and delivery in natural language, vary the pace between lines, add purposeful pauses, emphasize key phrases, and use controls such as style, speed, whispering, and intensity. Shorter script blocks also make it easier to shape each moment.

Can I use creative direction and inline voice controls?

Yes. Depending on the selected voice, you can combine a plain-language performance brief with line-level controls for style, speed, pitch, volume, emphasis, whispering, breathing, pauses, and intensity. Product names, acronyms, and technical terms can also be saved in a pronunciation library and reused across the project.

Can I use multiple voices in one project?

Yes. Assign voices to narrators, instructors, hosts, guests, and story characters, then keep those character choices consistent across sections. Dialogue workflows make it easy to reorder and direct multi-speaker scripts line by line.

Can I create training videos, courses, and narrated presentations?

Yes. Generate consistent narration by lesson or section, turn PowerPoint decks into narrated videos, create captions and transcripts for an LMS, and localize the finished course into 100+ languages from the same project workflow.

What can I export from a voice project?

Audio exports as a mastered MP3 at 192 kbps; Dubbing Studio also delivers a WAV alongside it. Depending on the studio, the same project can produce MP4 video, SRT or WebVTT subtitles, transcripts, audiograms, videograms, and shareable links.

Does Unmixr support AI voice cloning?

Yes. Unmixr supports consent-first voice cloning for approved voices, including multilingual use cases. Only clone a voice you own or have explicit permission to use, and follow the consent flow in the product.

Can I update one line without regenerating the whole recording?

Yes. Scripts are organized into editable blocks, so you can change and regenerate only the line or section that needs a retake. Approved audio, character assignments, pronunciation rules, and earlier versions can remain intact.

Still have a question?

If you have a specific use case or need help configuring a custom workflow, our team is always here for you. We would love to listen to your needs!