ElevenLabs: turn a short script into clear narration
Choose a voice, test pronunciation and pauses, then export a short narration and check it against your video.
The external player loads when you choose to watch. The source link is always available.
Steps and prompts
Follow the steps, example prompts and result checks below.
Your deliverable
Narrate a short coffee video with three shots: grinding, pouring and the finished cup. Export an MP3 or WAV and align the sentences with those shots. Let the actual recording determine the timing.
Follow the five steps
-
Prepare words that are easy to read aloud
Paste the three sentences below with their punctuation and paragraph breaks. Expand abbreviations and units into the words you want listeners to hear. Read the draft aloud once.
Check: Each number and unit has one intended pronunciation.
-
Compare voices with the same line
Open Text to Speech and use Voices to preview candidates. Try the opening sentence in two voices, then choose one that suits the quiet coffee footage. Note its name.
Check: Choose by the sound of your actual language and text, not only the voice preview.
-
Generate the complete short script
Choose an available model that supports your target language and select Generate Speech (called Generate in some views). Save this baseline as A and note the voice, model and text.
Check: All three sentences are present, with correct numbers and no missing words.
-
Revise one problem at a time
If the delivery is rushed, shorten a sentence or adjust punctuation and generate version B. Use speed controls or audio tags only where the chosen model supports them.
Check: Keep the version with natural pauses and complete sentence endings.
-
Download and align with the pictures
Open the generation history on the right; on a narrow screen, use the history icon above the generation button. Choose the download icon and MP3 or WAV. Place the audio on your editing timeline.
Check: Watch the whole sequence to check volume, sentence-to-shot alignment and the end of the recording.
Try this original example
These examples were written by FindGoodAI for practice; adapt them to your own material.
An original short narration script
This weekend, give yourself a moment to slow down.
Fifteen grams of coffee and two hundred and twenty-five milliliters of hot water let the aroma unfold.
As the final drop falls, your first cup of the day is ready.
When the result needs work
A number or abbreviation sounds wrong
Write the complete spoken form and replay the sentence containing it.
The voice seems to change between sentences
Keep the voice, model and settings consistent. Generate the short script together and inspect any joins.
A pause tag does nothing
Check the selected model’s documentation. Start with short sentences and punctuation instead of copying tags intended for another model.
The narration is longer than the footage
Shorten the script or allow the relevant shot more time, then check the whole sequence again instead of forcing excessive speed.
Before you finish
- Names, numbers and units are pronounced correctly.
- Pauses suit the shot changes.
- The narration is not unnaturally fast to fit a preset duration.
- The script, model, voice name and final audio are saved.
Match the video to the current interface
The official Text to Speech guide embeds this video. Model names and controls can differ from the recorded interface; use the choices currently available in your account. Generate and Generate Speech refer to the generation action in the official guide. The “111 Seconds” in the video title is not presented here as a verified runtime.
Video by ElevenLabs. The written steps, examples and checklist are an editorial practice guide based on the linked official documentation. Practice time is an estimate; generation queues may add time.


