AI voice-over
Turn text into speech in LUVI: pick a voice-over model and a voice, set the stability and language, and generate.
What this video covers
There are three things you can do with sound in LUVI: turn text into speech, make music, and transcribe a recording. This part covers voice-over.
Audio is a type of model in the workspace. Press Choose Model and open the Audio tab.
Video models that make their own sound are listed here too, and the audio models come last. The quickest way is to search for the model by name.
Several models do voice-over: ElevenLabs, Gemini, MiniMax Speech and xAI. We're choosing ElevenLabs v3.
For voice-over, the prompt box holds the script itself: whatever you type is read out word for word.
The Voice menu lists ready-made voices. The icon shows whether a voice is female or male, and the brackets show its accent.
A low Stability gives a more expressive read; a high one keeps it steady and even.
In the Language box, type a language code — EN for English.
The price depends on how long the text is, and the estimate updates as you type. Let's press Generate.
A few seconds later, the voice-over is ready. Press play to hear it.
The voice-over is saved to your library as an audio file. In the next part, we'll make music.
More from Going further
All tutorials
1:43Tutorial 14Editing your images
Edit an image with an editing model: put it in the reference slot, say what should change and what should stay, keep the aspect ratio and combine several images.
1:31Tutorial 15aYour first workflow
Build your first chain in Workflow: add nodes, pick models, wire an image into a video, check the price and run it.
1:50Tutorial 15bBatch generation in Workflow
Run one line with many inputs using Batch, Dice and Stack, and see how each one affects the price.
1:49Tutorial 15cWorkflow operators
Shape the text before it reaches the Generator with the Merge, Caption, Style and Pose operators.
1:49Tutorial 15dManaging workflow runs
The Preflight Check, running part of a graph, groups, cancelling, saving workflows and the panels on the right of the canvas.
1:26Tutorial 16bAI music
Make background music, or a full song with your own lyrics, in LUVI with MiniMax Music 2.6.
1:06Tutorial 16cTranscribing audio
Turn a recording into text with xAI STT v1, with speaker labels and filler words when you need them.
2:06Tutorial 17Talking avatars and lip sync
Make a portrait talk from an image and an audio file, or give an existing video a new voice with lip-sync models.
1:23Tutorial 18aGenerating 3D models
Turn an image or a description into a 3D model with Hunyuan 3D or HI3D, and see how extra views and parameters shape the result.
1:17Tutorial 18bThe 3D viewer and downloads
Inspect your 3D model in the viewer with lenses, lighting, wireframe and texture views, then download it as a GLB.