Skip to content
Synthesia logo

Synthesia

AI presenter videos from a script, built for training and explainers.

VideoMarketingEducation

How it works / How to use

AI presenter videos from a script, built for training and explainers. Write a script, choose a stock or personal AI avatar, pick language and voice, then generate a presenter-led video. Minutes and credits are plan-capped; dubbing and voice features draw from the same credit pool. Synthesia supports many languages and voices (exact counts vary by plan), stock and personal avatars, optional voice cloning with a custom personal avatar, and API access on Creator and Enterprise plans.

  1. Create a video from a template or blank project in Synthesia Studio.
  2. Paste or write your script, split into scenes/slides, and assign an avatar and voice per scene if needed.
  3. Add visuals (text, media, backgrounds) and preview pronunciation; adjust breaks and emphasis.
  4. Generate the video, review the output, then export, share, or request dubbing into another language from your credits.

Training and onboarding videos

One topic per video or chapter. Lead each scene with a clear spoken line for the avatar, then support with on-screen bullets or b-roll—not long paragraphs the presenter reads verbatim without breaks.

What to provide

  • Learning objectives and audience role
  • Script broken into short scenes (roughly 30–90 seconds each)
  • Tone (formal training vs conversational onboarding) and any compliance language that must appear verbatim

Details that improve the result

  • Short scenes are easier to regenerate if one line sounds wrong
  • Spell acronyms and product names phonetically if needed
  • Say whether on-screen text should mirror the script or stay minimal

Example prompt

3-scene onboarding video for new sales hires. Scene 1: welcome and goals (45s). Scene 2: CRM login steps—script must say the product name 'Northwind CRM' exactly. Scene 3: quiz reminder. Professional tone, medium-paced voice, 16:9, stock avatar in business casual.

If the first output is not good

  • “Regenerate scene 2 only; slow pacing on the login steps.”
  • “Add on-screen bullets matching scene 2, max 5 words each.”
  • “Swap to a warmer voice; keep the same script.”

Common mistakes

  • One 10-minute monologue in a single scene
  • Unpronounceable jargon with no phonetic hint
  • Regenerating the whole video to fix one misread word

Product explainers and how-to clips

Open with outcome (“By the end you will…”), then one step per scene. Keep spoken lines concise; put detail on screen when possible.

What to provide

  • The single job the viewer should complete
  • Step-by-step script with action verbs
  • Screenshots or screen recordings to embed, if used

Details that improve the result

  • Match avatar gestures to simple actions, not complex UI clicks unless shown on screen
  • Call out version-specific UI labels exactly
  • Keep explainer length aligned with plan minutes/credits

Example prompt

90-second explainer: export a report from Analytics. Scenes: (1) navigate to Reports, (2) choose date range, (3) download CSV. Script uses exact button labels. Friendly explainer tone, 16:9, light office background.

If the first output is not good

  • “Insert screenshot on scene 2 left half; avatar on right.”
  • “Shorten scene 1 to 15 seconds.”
  • “Regenerate voice only for scene 3.”

Common mistakes

  • Describing UI steps that are not visible on screen
  • Cramming feature marketing into a how-to script
  • Ignoring credit/minute caps on longer cuts

Multilingual videos and AI dubbing

Finalize the primary-language video first, then dub or recreate scenes per language. AI dubbing draws from the shared credit pool—budget languages up front.

What to provide

  • Source video or script in the primary language
  • Target languages from Synthesia’s supported set (check in-product for your plan)
  • Whether on-screen text also needs translation

Details that improve the result

  • Keep idioms out of the base script when you plan many locales
  • Review lip-sync sensitive scenes after dubbing
  • Note that dubbing and generation share plan credits

Example prompt

Base English training video (3 scenes, 4 minutes total). Dub to European Spanish and French for EU staff. Keep on-screen product name in English; translate spoken instructions only.

If the first output is not good

  • “Redub scene 2 in French with slower pace.”
  • “Replace idiomatic phrase in English source before re-dubbing.”
  • “Add subtitles file export if available on plan.”

Common mistakes

  • Dubbing before the English script is final
  • Untranslatable jokes or culture-specific examples
  • Assuming dubbing is unlimited on lower plans

Stock avatar presenter videos

Pick a stock avatar that fits audience and tone, then write to the avatar’s natural pacing. Stock avatar counts vary by plan—confirm availability before scripting multiple cast members.

What to provide

  • Script and desired presenter style (business, casual, energetic)
  • Framing (torso, standing, seated) if the template allows choice
  • Background and brand colors

Details that improve the result

  • One primary avatar per video keeps continuity
  • Match background to brand without busy motion behind the face
  • Preview the first scene before generating the full video

Example prompt

Internal policy update, 2 minutes. Calm professional stock avatar, neutral studio background, navy lower-third bar. Script is plain language, no jokes, 16:9.

If the first output is not good

  • “Try a different stock avatar with similar tone.”
  • “Move avatar to left third; add bullet list on right.”
  • “Reduce background contrast behind the face.”

Common mistakes

  • Switching avatars mid-video without narrative reason
  • Busy backgrounds that compete with the presenter
  • Assuming every stock avatar supports every language equally well

Personal avatar with voice cloning

Voice cloning can be paired with a custom personal avatar on supported plans. Record clean sample audio if required, then test short lines before generating long training content.

What to provide

  • Approved personal avatar setup per plan rules
  • Voice clone recording or approval workflow, if offered on your plan
  • Script that sounds natural when spoken in your voice

Details that improve the result

  • Write conversational sentences, not bullet-speak
  • Avoid excessive numbers and abbreviations in the first test clip
  • Confirm personal avatar and clone entitlements on your plan

Example prompt

CEO update video, 60 seconds, using approved personal avatar and cloned voice. Script: three short paragraphs, contractions allowed, no legal disclaimers beyond the provided compliance sentence.

If the first output is not good

  • “Regenerate with shorter sentences in paragraph 2.”
  • “Add a pause marker before the compliance line.”
  • “Lower energy; keep clone voice.”

Common mistakes

  • Using voice clone without plan approval or consent workflow
  • Reading dense legal text in a casual cloned voice
  • Long first generation before testing a 10-second sample

Template-based updates at scale

Lock design in the template; swap only approved variables (name, role, metric, language). Batch similar videos instead of one-off edits when content repeats.

What to provide

  • Master template with fixed layout and variable fields
  • Data source for names, titles, metrics, or locales
  • Rules for what must not change (logo, disclaimer, colors)

Details that improve the result

  • Keep variable slots obvious in the script labels
  • Validate numbers and names before batch generate
  • Track credits when producing many variants

Example prompt

Monthly manager update template: slots {manager_name}, {team_win}, {metric_value}. Generate 12 versions from a spreadsheet. Fixed closing compliance line must appear verbatim in every video.

If the first output is not good

  • “Regenerate only rows where metric_value changed.”
  • “Add pronunciation hint for manager_name in row 7.”
  • “Shorten template intro by 10 seconds for all variants.”

Common mistakes

  • Letting variable text overflow designed text boxes
  • Batch generating before one template proof is approved
  • Inventing metrics not present in the source sheet

API-driven video generation

API access is included on Creator and Enterprise plans. Treat scripts and template variables as your contract; validate render status before publishing URLs downstream.

What to provide

  • Creator or Enterprise plan with API access
  • Script/template identifiers and avatar/voice IDs supported by the API
  • Webhook or polling for render completion

Details that improve the result

  • Start with one template and one language before scaling
  • Store Synthesia asset IDs rather than hard-coding names
  • Handle render failures with retries and operator alerts

Example prompt

POST generate from template_id=onboarding_v3 with variables { employee_name, start_date } and avatar_id=stock_042, voice=en-US. Poll until complete, then save MP4 URL to HRIS record.

If the first output is not good

  • “Fallback to stock avatar if personal avatar unavailable.”
  • “Reject generate if start_date missing.”
  • “Notify Slack #video-ops on render failure.”

Common mistakes

  • Assuming API access on Starter when it requires Creator or Enterprise
  • No timeout or failure handling on long renders
  • Changing template variables without versioning consumers

How to prompt

Synthesia is script-first: you write what the avatar should say, then shape scenes, voice, and visuals. Prompt-like planning means clear spoken lines, scene boundaries, pronunciation hints, and on-screen support—not open-ended chat. Budget minutes and credits, especially for dubbing and batch runs.

Video: 2-scene GDPR refresher, 16:9, professional tone.
Scene 1 (40s): define personal data—use exact phrase 'personal data' twice.
Scene 2 (50s): three employee actions as numbered spoken lines.
Avatar: stock business casual. Voice: en-GB, medium pace.
On-screen: max 5-word bullets per scene.

Write for the ear, not the page

Short sentences, contractions where appropriate, and pauses between ideas sound more natural in presenter videos.

Example

Say: 'Open Settings. Choose Privacy. Turn on two-factor auth.' Not: 'Users should navigate to the Settings area in order to configure privacy-related authentication enhancements.'

Split by scene early

Scene breaks map to regeneration boundaries and keep credits focused when one section needs a fix.

Example

Scene 1 welcome (30s). Scene 2 demo (45s). Scene 3 recap (20s). Each scene gets its own slide and avatar block.

Spell tricky words and brands

Product names, acronyms, and uncommon terms may need phonetic hints or simplified phrasing.

Example

Script line: 'Open Northwind CRM' — note pronunciation: 'NORTH-wind C-R-M'. Avoid 'NwCRM' in speech.

Pair speech with on-screen structure

Put lists, steps, and labels on screen; let the avatar introduce and transition, not read dense bullets verbatim.

Example

Avatar: 'Here are three steps.' On-screen numbered list appears while the avatar summarizes each step in one sentence.

Plan languages and credits together

Multiple languages and dubbing share your credit pool—finalize the primary script before multiplying locales.

Example

Lock English script v3, generate once, then dub to es-ES and fr-FR only. Do not dub draft v2.

Best output tips

Script in scenes, not one block

Break training and explainer content into scenes with clear durations. Scene boundaries are your best unit for fixes and credit control.

Optimize for spoken delivery

Use short sentences, active verbs, and natural transitions. If it sounds stiff when read aloud, the avatar will sound stiff too.

Lock compliance copy verbatim

Legal, safety, and policy lines should be pasted exactly as approved. Mark them in the script so they are not paraphrased during edits.

Choose avatar and voice to match audience

Stock avatars cover many presenter styles; personal avatars and voice clones are plan-gated. Pick one primary presenter per video for continuity.

Support speech with visuals

Show steps, diagrams, and labels on screen. Presenters work best as guides, not as readers of dense paragraphs.

Finalize before dubbing

Synthesia supports many languages and voices; AI dubbing draws from shared credits. Finish the primary-language script and timing before generating alternate languages.

Watch minutes and credits

Presenter videos are plan-capped by minutes and credits. Estimate runtime early, especially for batches and dubs.

Test clones and personal avatars small

Generate a short sample with voice cloning and personal avatars before committing to long-form content. Adjust pacing and sentence length based on that sample.

Use templates for repeatable formats

Monthly updates, personalized welcomes, and regional variants benefit from fixed layouts with controlled variable fields.

Integrate via API deliberately

Creator and Enterprise include API access. Version templates, validate required variables, and handle render failures before connecting HR, LMS, or marketing systems.

  • Synthesia is script-driven: write spoken lines first, then scenes, avatar, voice, and visuals.
  • Keep scenes short so you can regenerate one section without rebuilding the whole video.
  • Preview pronunciation on acronyms, names, and product terms before full generation.
  • Stock and personal avatar availability varies by plan—confirm before scripting multiple presenters.
  • Voice cloning pairs with custom personal avatars on supported plans; test a short clip first.
  • AI dubbing and generation draw from the same credit pool; budget multilingual work up front.
  • Use templates and variable slots when you need many similar videos (updates, localized variants).
  • API access is on Creator and Enterprise plans—validate plan before building integrations.
  • Put detailed steps on screen; let the avatar summarize instead of reading long bullet lists.
  • Review generated video for lip-sync, pacing, and compliance language before publishing.

Try this AI

Try Synthesia

Product Details

Pricing, features, limits and latest updates

Synthesia

AI presenter videos from a script, built for training and explainers.

Free / Paid · Free

Pricing Plans

Basic

Free

Starter

$29/ Monthly

Creator

$89/ Monthly

Enterprise

Custom

Key Features

Video Generation

Creates presenter videos from a script; minutes and credits are plan-capped.

Voice

Includes 160+ languages and voices plus AI dubbing from the shared credit pool.

API

API access is included on Creator and Enterprise plans.

Voice Cloning

Voice cloning can be paired with a custom personal avatar.

AI Avatars

Stock and personal AI avatars present the scripted video; counts vary by plan.

Limits

  • Video generation: 2 credits/second. Presenter video generation consumes two credits per second on self-serve plans (Synthesia credits help).
  • Assistant: Synthesia Assistant does not consume credits; final video generation does (Synthesia Assistant help).
  • API access: API access requires Creator plan or above (Synthesia API docs).

Ideas / Prompt experiences

Share a prompt that worked for you. Username and email are shown with your submission. External links are not allowed.

Example prompt

Assistant → script-first explainer

Prompt

Assistant prompt: Create a 90-second training video explaining two-factor authentication for new employees. Tone: friendly and practical. Include one on-screen checklist scene and a plain-language recap.

Short explanation

Synthesia Assistant helps draft outline and script; final video generation consumes credits separately.

Example prompt

Scene-based presenter script

Prompt

Scene 1 (15s): Welcome and goal. Scene 2 (45s): Step-by-step setup with on-screen labels. Scene 3 (20s): Common mistakes and recap. Voice: clear presenter, neutral accent. No jargon without definition.

Short explanation

Script-first workflows work best with short scenes and visuals supporting speech, not dense paragraphs.

Share your experience

Required fields are marked. Variation, result, and explanation are optional.

Pika logo

Pika

Explore

AI video platform for text-to-video and image-to-video generation with credit-based plans and tools including Pikascenes, Pikadditions, Pikaswaps, and Pikatwists.

VideoFree / Paid

CapCut logo

CapCut

Explore

Video editing with AI captions, effects, and generative clips.

VideoFree / Paid

Adobe Podcast logo

AI audio recording and editing on the web, including Enhance Speech for noise removal and Studio for browser-based podcast production.

Video