---
updatedAt: 2026-09-17T12:04:58.000Z
---

Fetch the complete documentation index at: https://docs.synthesia.io/llms.txt. Use this file to discover all available pages before exploring further. Append .md to any documentation page URL to get its markdown version.

# Script box

Add a your script to the script box to transform your text to speech and adjust the pronunciation, pacing, and emphasis of the speech as needed.

Create your script by either typing or copying and pasting text into the script box.

The maximum duration of the script for each scene is 5 minutes.

Your script can be in any of the [languages that Synthesia supports](https://docs.synthesia.io/docs/supported-languages). You can use multiple languages in a single scene and/or video by applying different voices to paragraphs of your script in the script box—each paragraph (hard return) can have its own voice, allowing you to assign different speakers, voices, and languages to each paragraph if needed. Synthesia automatically selects a voice that matches the detected language of your script, but you can always change the voice from the voice picker.

Once you've added your script, you can control the speaker and the voice that will be used by selecting the speaker pill on the left side of the script.&#x20;

![Speaker pill](https://files.readme.io/a59cae715df148ea7ffb321bf87c42aee8b19660f41a0fe3bb7f04acbc17f7a2-ezgif.com-optimize_6.gif)

When you create line breaks by pressing Enter on your keyboard, you create new paragraphs. Each paragraph has its own speaker pill, so you can adjust the speaker, voice, voice speed, and regenerate speech for it independently of the others. This allows you to create dialogue (by setting one or more different speakers throughout the script) and otherwise fine tune each paragraph as needed.&#x20;

<Callout icon="📌" theme="default">
  ### Note:

  If you're looking to edit the transcript for a dubbed video, check out the [dubbing](https://docs.synthesia.io/docs/video-dubbing#proofread-and-correct-the-transcript) page.
</Callout>

# Speaker

If you have more than one speaker or avatar in your scene, select the left side of the speaker pill to assign the desired speaker to that line:

<Image src="https://files.readme.io/be5c0e379a19c8667beb7818796acd38802b2853cd2ada1ceb38a52f7596badb-Screenshot_2026-06-04_at_4.35.43_PM.png" alt="Selecting a different speaker" align="center" width="350px" />

In this modal:

* Selecting the eye icon toggles the visibility of the avatar in the scene. Disable it if you just want the voiceover.
* Selecting `+ New speaker` allows you to pick a new avatar to use as the speaker for that paragraph (and adds it to your scene).

# Voice settings

<Image src="https://files.readme.io/821c2aaa9f38ff00b728d467aed574ab4ae1a3fc35ffe37f45bb902fb4bbfbef-image.png" alt="Voice settings panel" align="center" width="350px" />

From the `Voice settings` panel you can preview the current voice, adjust speed and `Variant`, or select `Change voice` to open the voice picker and choose a different voice or language.

## Select a different voice for a section of the script

Select `Change voice` to open the voice picker. By default, it's filtered to the language and accent currently in use:

<Image src="https://files.readme.io/945e39c389c8899cab77714cd802b1eb2dfb0441765f7b8ac896bb789ec99ea7-image.png" align="center" width="550px" />

1. Browse `Featured voices`—Synthesia's recommended picks—at the top of the picker, or go to `All voices` and choose a tab:
   1. &#x20;`From Synthesia` for Synthesia's built-in voice catalog
   2. `My voices` for your own voice clones (available if you've created one using Synthesia's voice cloning feature or a [personal avatar](https://docs.synthesia.io/docs/personal-avatars))
   3. `Shared with me` for voices a teammate has [shared with you](https://docs.synthesia.io/docs/custom-voices#voice-sharing).&#x20;
2. Optionally, search for a voice by name, filter the list by language or accent, or use the sort control to reorder it. To assign a different language to this paragraph, filter by that language and pick a voice from the results—the [languages available](https://docs.synthesia.io/docs/supported-languages) are the same ones supported elsewhere in Synthesia.
3. Preview a voice by selecting the play button on its row, or select the row itself to load it into the preview bar at the bottom of the picker.
4. Select the `+` button on a voice's row, or select `Swap voice` in the preview bar, to apply that voice.

<Callout icon="📌" theme="default">
  ### Note:

  If a voice you expect to see is missing, open the filter options and turn off the `Quality` toggle—when it's on, it hides voices Synthesia doesn't recommend:

  ![](https://files.readme.io/90071fd0e79f35e6b7af0f7b9686af3d914d8e842f18080cfa37de5f2060ded4-Group_2085663935.png)
</Callout>

## Change the voice variant

A voice can have several variants—different styles of the same voice, each powered by different underlying speech technology. Every voice defaults to the `Auto` variant, which always uses the latest and highest-quality technology available for that voice. Most of the time you won't need to change this.

If you're sensitive to small changes in how a voice sounds, you can lock a voice to a specific variant:

1. Select the voice, then open `Voice settings`.
2. Open the `Variant` dropdown.
3. Select a variant from the list. Each option has an info (ⓘ) icon with details about that variant.

<Image src="https://files.readme.io/47789f5976aad7a4218cf780fbd8c031e149ab4719c97d2903f8a6fd38e58b07-image.png" alt="Voice variants" align="center" width="350px" />

<Callout icon="📌" theme="default">
  ### Note:

  An avatar's voice uses a single variant for the whole video—you can't switch variants between scenes.
</Callout>

### Remastered voices and pronunciation controls

Remastered voices use the `Synthesia Express-Voice` variant by default and unlock [pronunciation controls](#pronunciation). They may sound slightly different from the previous version.

If you notice that a voice you've been using has changed, you can return to the original sound: open the `Variant` dropdown and select `ElevenLabs Turbo v2`. Note that this disables pronunciation controls for that voice. Because remastered voices use `Express-Voice`, previewing them can take a few seconds longer than the older variant did.

## Voice speed controls

Use voice speed controls to fine-tune how fast a voice speaks, so you can match the pacing to your content and audience. You can apply speed changes per paragraph in your script, or across the whole video with one setting. This gives you flexibility to slow down dense, instructional content while keeping other parts of the video more dynamic. Voice speed controls for specific sections of a single sentence or paragraph aren't available.

<Image src="https://files.readme.io/3e382a02ff242fc83133c16a695ba1c4505b4b025f3d4fd5f8539487a73b56f2-Screenshot_2026-01-21_at_2.46.43_PM.png" alt="Voice speed control" align="center" />

**To adjust the voice speed for a paragraph:**

1. Hover over the speaker pill on the left side of that paragraph.
2. Select the speed icon on the right side of the speaker pill.
3. Use the slider or directly input your desired voice speed. The numeric values (for example, 1.2×) are approximate guides for speed, not exact percentages of the default speed.
   * Minimum speed: 0.8× (slower, more deliberate delivery)
   * Maximum speed: 1.2× (faster, more energetic delivery)

**To standardize the voice speed across the whole video:**

Select the `Change all` button in the toast notification at the bottom of the screen after adjusting the voice speed for a specific paragraph.

<Image src="https://files.readme.io/fd06be29ff97a80a2d249beb594b1f11197daf98caffd5732c99758a4c68c172-image_381.png" alt="Change all voice speed" align="center" />

We recommend previewing your video after adjusting the speed to ensure it still sounds natural and remains easy to follow.

## Speech regeneration

Use speech regeneration to get an alternative take of the same script with the same voice, without changing your script, settings, or visuals. When you regenerate speech for a line, paragraph, or scene, Synthesia keeps your script, language, and voice exactly the same while producing a new audio take with subtle differences in timing, intonation, rhythm, and emphasis on certain words.

To regenerate audio for the entire scene, select the `Regenerate` (arrow) icon to the right of the `Preview scene` button:

![Regenerate audio for this scene](https://files.readme.io/5197c042747f6fa02fe27b7cb05a0f41b60907a85d5a2a715ad6841c19e6e4c3-Screenshot_2026-06-04_at_4.04.58_PM.png)

To regenerate audio for a paragraph, either:

1. Select a paragraph and right-click to access the context menu, then select the `Regenerate` option.
2. Select a paragraph and select the `Regenerate` (arrow) icon to the right of the `Preview` button:

![Regenerate audio for this paragraph](https://files.readme.io/1e45365231359113cfaa7f970cdc6075cc23ca2af3b13310479267e30f486682-Screenshot_2026-03-16_at_11.54.10_AM.png)

You can regenerate multiple times and keep the take that works best. If you change your mind, you can revert back to a previous speech version with the undo button.

This is especially helpful when a word sounds slightly off but doesn't need a pronunciation fix, when the delivery feels too flat in one spot, or when you want a version that better matches the energy of a specific scene.

Synthesia's AI voices are non-deterministic, which means the same script and voice can sound slightly different each time you generate audio. Speech regeneration lets you take advantage of that to quickly find a version that sounds more natural for your use case.

# Additional script actions

Right-click anywhere in the script box to display all script actions available. Highlighting text before right-clicking specifies the script text that the actions are applied to:

<Image src="https://files.readme.io/680dc3a55b7b3cd07b6b084c5dd359423d2f286167ee2a25939167a0f098f7ce-Screenshot_2026-03-16_at_11.54.41_AM.png" alt="Script menu" align="center" width="250px" />

Highlighting a portion of your script also reveals a menu above it, with quick access to several script actions:

<Image src="https://files.readme.io/f594b1ea546cf615e8d31459fcbef30c36a13f517a507d5499b499fce25c6245-image.png" align="center" />

* `Edit with AI`
* `Pronunciation`
* `Add translation`
* `Preview`
* The regenerate icon—see [Speech regeneration](#speech-regeneration).
* The `•••` (more) icon—opens the same options available from [right-clicking in the script box](#additional-script-actions).

See the sections below for what each of these does.

## Edit with AI

Use Edit with AI to rewrite, lengthen, shorten, adjust the tone of, summarize, or translate part of your script, without leaving the script box.

![](https://files.readme.io/5ab790f6bbb0ba434cf27a230637b1798896bfb59ec07cd46441409e613a3df1-image.png)

**To edit part of your script with AI:**

1. Highlight the portion of the script you want to edit.
2. Select `Edit with AI` from the menu that appears above the highlighted text. This opens the AI Video Assistant panel.
3. Either:
   * Type your own instruction instead—for example, ask it to translate the highlighted text into another language, or describe any other edit you want.
   * Select one of the preset options: `Rewrite`, `Make it longer`, `Make it shorter`, `Change tone`, or `Summarize`.
4. Select the arrow button to submit and preview the result.
5. If you're happy with it, select `Replace` to swap in the new text. Otherwise, refine your instruction and try again.

## Pause

Use pauses to pause the script for a defined period of time. Use them to adjust the pacing of your script and the emotional tone of the voice being used.

<Image src="https://files.readme.io/01b66b5b1dc69ec7e55a22d1ce3f5b7e8f7c4f0de32fd89696b51579ed1b8a86-image.png" alt="Pauses" align="center" width="250px" />

**To add a pause to your script:**

1. Position your text cursor where you want to add a pause, and either:
   1. Right-click and select the `Pause` option from the context menu.
   2. Select the `Pause` button at the top-left of the script box.

      <Image src="https://files.readme.io/b22730d0d077b4ad132a77e35b545f0e619f56f7b3310fdfe0a477f988f455c8-Screenshot_2024-11-29_at_1.01.40_PM.png" alt="Pause button" width="350px" />

2. Adjust the pause duration to a value between 0.1s and 99s (the default is 1 second).

Pauses can be dragged and dropped to change their position within the script. You can also highlight a pause to copy and paste elsewhere in your script.

## Pronunciation

Use the pronunciation feature to adjust the pronunciation of words and phrases in the script—useful for brand names, technical terms, and acronyms.

<Image src="https://files.readme.io/c707b913a53b34c04499b7e196c80039f6161515f76974e464af3b20e96c4e66-Screenshot_2026-09-17_at_7.24.38_AM.png" alt="Pronunciation" align="center" width="550px" />

**To modify the pronunciation of a word or phrase:**

1. Highlight the text you'd like to adjust the pronunciation for, and either:
   1. Select `Pronunciation` in the modal that appears above the highlighted text.
   2. Right-click and select the `Pronunciation` option from the context menu.
   3. Select `Pronunciation` at the top-left of the script box.
2. Choose how to set the pronunciation:

   * **Type Pronunciation**: Enter a phonetic spelling that sounds right when read naturally (for example, "sin-THEE-zhuh" for Synthesia).
   * Select `Record yourself` and say the word aloud—Synthesia captures your pronunciation and uses it as the reference.

   <Callout icon="👍" theme="okay">
     ### Tips for defining a pronunciation

     Select `Tips` for guidance:

     - Plain language works for most words—use phonetic spelling to stress a syllable, or IPA if you know the exact sounds (for example, "tomato" as "toh-MAY-toh", or "data" as "DAY-tuh").&#x20;
     - For numbers, symbols, and abbreviations, spell out exactly what the voice should say, since it reads your text literally (for example, "222" as "two hundred and twenty-two", "mg" as "milligrams", or "St" as "street" or "saint").

       ![](https://files.readme.io/aa5a7b82be07f77c0d198f03d48dc8e70fccc6cdd575e96666ddddc3f2c8c846-Screenshot_2026-09-17_at_7.24.50_AM.png)
   </Callout>
3. Select `Next`. Synthesia analyzes what you entered and offers two versions of the pronunciation to preview and choose from:
   * **Option 1** uses IPA (International Phonetic Alphabet) and is best for fine-tuning how a word sounds phonetically.
   * **Option 2** uses a respelling that replaces the text before it reaches the voice model, making it more consistent for numbers, dates, symbols, and abbreviations.

     ![](https://files.readme.io/1b7b072ed440413b00860f3d216f142e4bded271a4898f89175b5e1132ff0905-Screenshot_2026-09-17_at_7.25.32_AM.png)
4. Preview each option, then select `Use` to save the one you prefer, or `Back` to adjust your input and try again.
5. Optionally:
   * Apply the pronunciation to all instances of that word or phrase throughout your video by selecting `Apply to all` in the toast notification that pops up at the bottom of the editor after confirming your pronunciation preference. This doesn't automatically apply the pronunciation to additional instances of that word or phrase that you add to your script afterward.
   * Add the pronunciation to your workspace's [Glossary](https://docs.synthesia.io/docs/translation-glossary#pronunciation)—the option you chose (IPA or respelling) is saved with it, so it's applied automatically and consistently the next time anyone in your workspace types that word.

![Change all pronunciation](https://files.readme.io/c89be054fc6343433e96d2d30fa45e9863aaa18090e2158170a04ff82cfd40a1-Group_2085662921_2.png)

By default, a Glossary pronunciation applies across every voice in a language, not just the one you set it on—you can narrow this to a specific voice or locale (for example, British English only) from the [Glossary](https://docs.synthesia.io/docs/translation-glossary) page.

<Callout icon="📌" theme="default">
  ### Note:

  Pronunciation reliability varies by voice and language—voices created by Synthesia tend to be the most consistent, while third-party voices and some older voices in non-English languages may be less reliable. If you're on a voice with limited reliability, a notice appears in the pronunciation menu with the option to switch to a more reliable one.

  If a word is still mispronounced after saving a pronunciation, try [regenerating the audio](#speech-regeneration) for that scene.
</Callout>

## Add a translation

**To add a translation:**

1. Highlight the term in the script and select `Add translation`.
2. From the panel, you can toggle `Don't translate`, view and select existing translations (if available), or select `Manage translations` to open the full [Glossary](https://docs.synthesia.io/docs/translation-glossary) page.

![](https://files.readme.io/6ccfc5b14d2e649f414af44e3f81c3bcae2b5cdc99849988c4bdf69a43d70ecc-ScreenRecording2026-08-24at12.59.50PM-ezgif.com-crop.gif)

## Previewing the scene

You can preview the entire scene or just a portion of it in a few different ways.

### Preview the entire scene

<Image src="https://files.readme.io/63d5bffdfe690e0ed2f293928bddb2c16a8538c85d8c7ccfce0acfd98ed9be59-image.png" alt="Script preview button" align="center" width="450px" />

Select the `Preview scene` button (play icon) at the top left of the script box to launch a scene preview. You can use the playhead/seeker in the timeline to set a starting point for the scene preview.

Launching a scene preview:

* Launches a voice-over of the script.
* Animates all assets from the scene.
* Plays the music selected for the scene.
* Previews the transition to the next scene.
* Doesn't animate avatars on screen (their lips won't move).

### Generate an avatar preview

<Callout icon="📘" theme="info">
  ### Generating an avatar preview is available on Enterprise plans.
</Callout>

The preview above doesn't show the animation of the avatar unless the video has already been previously generated and the scene remains unchanged. To see and hear the avatar's full performance for a scene—without rendering the entire video—use Scene Preview.

![](https://files.readme.io/235cb1ecec5874c115b6c3ce5f2dea1f9d99e0afcebbf6905c56f8eef3b2354c-Group_2085663944.png)

**To generate an avatar preview:**

1. Select the dropdown arrow next to Preview, at the top of the script box, to open the *Generate preview* menu.
2. Choose one:
   1. Audio only: Previews the script instantly, without generating the avatar's performance.
   2. Generate avatar: Renders the avatar's full performance for this scene. This takes a couple of minutes, depending on the length of your scene.
3. While the avatar performance renders, its progress appears in the *Activity* panel, where you can select `Cancel` to stop it:

   ![](https://files.readme.io/68e2274c48ed231cca561cf3c4fe16799ecb8a4db8f51d46d61f02e3ade0727b-Screenshot_2026-09-17_at_7.46.13_AM.png)
4. Once the render finishes, the icon next to `Preview` turns green to show the avatar preview is ready:

   ![](https://files.readme.io/a0225b42e9f86838ff8c3b61f1de8dd86809e65a939e25c98b34d47c3ff0d272-Group_2085663945.png)
5. Click the `Preview` button to the left it to watch the performance, or regenerate it again for a new take:

   ![](https://files.readme.io/8631cf79c8fe7f776d4b91384ff0b0459172703bb132dc5acdfcabf9484ffc2b-ScreenRecording2026-09-17at7.51.50AM-ezgif.com-optimize.gif)

Synthesia reuses a scene's existing avatar performance for the preview when nothing has changed, so you don't need to regenerate it every time you reopen the editor. If you edit the script, voice, or framing for that scene, the existing avatar performance is discarded and you'll need to generate it again—each generation re-renders the avatar and voice together, so you can't refresh just one and keep the other.

Like any other render, generating an avatar preview is subject to Synthesia's usual content moderation.

### Preview the animations and scene transition

If you don't need to preview any audio and just want to make sure the visual design for your scene is on track, select and drag the playhead/seeker for the timeline at the top of the script box.

![Dragging the playhead to preview animations](https://files.readme.io/60361f161283eca64f16de3ff4b7190e644727133670032bdbb85f41767f0f35-ScreenRecording2025-03-19at2.10.14PM-ezgif.com-optimize.gif)

Dragging the playhead all the way to the end of the timeline previews the transition to the next scene.

### Preview part of the voiceover

![Previewing highlighted text](https://files.readme.io/9370377ff94c1e53871b29def3466982f1505950b023121bb8b92daa9843b1aa-Screenshot_2026-03-16_at_11.57.13_AM.png)

If you have a long script, preview only part of your text:

1. Highlight the portion of the script that you want to preview.
2. Then, either:
   1. Select `Preview` in the menu that appears above highlighted text.
   2. Right-click on the highlighted text and select `Play` from the context menu.

Previewing your voiceover this way doesn't launch a preview of the animations, music, or transition for the scene.

## Variables

Right-click in the script box and select `< > Variable` to insert a variable, for building templates you can use to create videos at scale. See [Add variables for programmatic video creation](https://docs.synthesia.io/docs/synthesia-templates#add-variables-for-programmatic-video-creation) on the Templates page.

## Audio file upload

<Callout icon="📘" theme="info">
  ### Audio file upload is an Enterprise plan feature.
</Callout>

![Uploading an audio file for a scene](https://files.readme.io/96d861be0d25b4d54024f1ea92231efe46ba9065aa67133ed45f8780ae810b22-image.png)

Instead of writing a script, you can upload an audio file with a voiceover to be spoken by the avatar.

You can upload a maximum of 5 minutes of audio per scene. Supported formats: `.mp3`, `.flac`, `.wav`, `.m4a`.

**To upload an audio file:**

1. Select upload in the script box.
2. Choose the desired file.
3. Select `Open` to upload the file.
4. Specify the language of the audio file.

It's not possible to change the volume or playback speed of uploaded audio.

## Script remediation guidance

Script remediation guidance gives you feedback on script issues before generating your video, so you can fix them upfront instead of waiting for a rejected submission.

This feature is off by default for Enterprise plans—[organization admins](https://docs.synthesia.io/docs/organization-settings#feature-settings) can enable it globally or per workspace, and [workspace admins](https://docs.synthesia.io/docs/workspace-settings#feature-settings) can enable it for their own workspace, from their respective settings. This setting can't be configured for non-Enterprise plans—script remediation guidance is always enabled there.

When you open the generate panel, Synthesia checks your script in the background. If it finds an issue, you'll see guidance on what to change before you select `Generate`. Fix the flagged content directly in the editor and resubmit without starting over.

<Callout icon="📌" theme="default">
  ### Note:

  This guidance is advisory only. You can still choose to generate your video without addressing it, and normal content moderation still applies as usual after you generate.
</Callout>

## Gestures

<Callout icon="📌" theme="default">
  ### Note:

  Gestures are a legacy feature only available for non-expressive avatars. For older avatar versions, expressiveness isn't automatic—you should manually add gestures in your script for the avatar to show emotion.
</Callout>

![Non-expressive avatar gestures](https://files.readme.io/e7e27a252c1c97cc155fcc9dc051c4a6fc0d421c94f27e2a71ccfa75c0f7ff51-image.png)

Available gestures are:

* **Nod**: The avatar nods down once.
* **Head Yes**: The avatar moves its head twice up and down.
* **Head No**: The avatar moves its head twice left and right.
* **Eyebrows Up**: The avatar raises its eyebrows.

# Sibling pages

* [Elements](https://docs.synthesia.io/docs/elements.md)
* [Timeline](https://docs.synthesia.io/docs/timeline.md)
* [Storyboard](https://docs.synthesia.io/docs/storyboard.md)
* [Animations & Effects](https://docs.synthesia.io/docs/animations-and-effects.md)

# What’s Next

* [Video Creation](https://docs.synthesia.io/docs/video-creation.md)
* [Video Edit page](https://docs.synthesia.io/docs/video-edit-page.md)
* [Supported languages](https://docs.synthesia.io/docs/supported-languages.md)