Explainer videos — the kind that walk through a concept with visuals and narration — are one of the most effective content types in online courses. They're also one of the most time-consuming to produce. Synthesia generates these videos from a script using AI avatars and text overlays, with no filming, no lighting setup, and no video editing. You write what you want explained, choose a presentation format, and Synthesia produces a polished video in minutes.
Quick Answer: How to Create Animated Explainer Videos Using Synthesia
- Write 2-Min Script: Execute this step in your course creation workflow.
- Select Avatar: Execute this step in your course creation workflow.
- Add Visual Aids: Execute this step in your course creation workflow.
- Generate & Review: Execute this step in your course creation workflow.
- Export: Execute this step in your course creation workflow.
What you’ll walk away with:
- Short, focused explainer videos for key concepts
- Supplementary content that complements your main lessons
- A library of reusable concept explanations
Why Synthesia for explainer videos
The traditional explainer video workflow is: write a script, set up camera and lighting, record (with multiple takes), edit the footage, add graphics and overlays, render, and export. Even for a three-minute video, that process takes a full afternoon. With Synthesia, you write the script and the platform handles everything else — avatar presentation, lip-sync, text overlays, scene transitions, and rendering.
For course creators, the most practical use cases are supplementary content that doesn't require your personal presence. Course introductions that welcome students to each module. Concept explainers that walk through a technical process or framework. Multi-language versions of content for international audiences. These are the videos where production quality matters and personal warmth matters less — a sweet spot for AI-generated content.
Synthesia's pricing has a free tier that gives you 10 minutes of video per month. The Starter plan ($29/month) adds more video minutes and better options. The Creator plan ($89/month) includes custom avatars and priority rendering. For most independent course creators, the free tier covers supplementary content needs.
How to create an AI explainer video
Write a clear, conversational script
Your script is the video. This is even more true with AI avatars than with real filming, because the avatar delivers exactly what you write — no improvisation, no natural recovery from awkward phrasing, no spontaneous emphasis. Write the way you'd explain the concept to one student. Short sentences. Plain words. Natural pauses between ideas (add a period or comma where you'd naturally pause if speaking).
Keep each explainer under three minutes. Research from Guo, Kim, and Rubin's analysis of 6.9 million edX video sessions shows that engagement drops significantly after six minutes even with real instructors. With AI presenters, shorter is better — viewers disengage faster when they detect the artificial delivery.
Choose your layout
Synthesia offers several layout options. For explainer videos, the most effective is the split-screen format: avatar on one side, slides or visuals on the other. This gives viewers something informational to look at while the avatar narrates. A full-screen talking avatar — with nothing else on screen — loses attention quickly because there's no visual content reinforcing the explanation.
Prepare your visual slides before entering Synthesia. Upload them as images (PNG or JPEG) that appear alongside the avatar at the right moments. Each slide should illustrate one key point from your script.
Select an avatar and voice
Synthesia offers 140+ stock avatars. For explainer content, choose one with a neutral, professional presentation style — something that fits the tone of educational content rather than marketing. Preview the voice options for your chosen avatar. Some voices handle technical terminology better than others. If your course is for a specific language audience, Synthesia supports 130+ languages with localized voices.
Add text overlays and transitions
Reinforce key points with text that appears on screen as the avatar mentions them. This dual-coding — hearing and seeing the same information — improves retention. Keep text overlays short: 3-7 words per overlay, matching the key term or concept being explained. Use Synthesia's scene transitions to break the video into clear sections, especially if your explanation has distinct stages.
Generate, review, and iterate
Generate the video and review with fresh eyes. Listen for pronunciation issues — AI voices sometimes struggle with technical terms, acronyms, or names. If "API" comes out wrong, spell it phonetically in the script: "ay-pee-eye." Check that text overlays appear at the right moments. If anything feels off, edit the script and regenerate — this is one of the real advantages of AI video over traditional filming.
Export and embed in your course
Download as MP4 at 1080p resolution. Upload to your course platform and embed in the relevant lesson. On Ruzuku, video uploads play directly in the lesson step — no external hosting required. Place the explainer video where it adds the most value: at the start of a module (as an overview), before a hands-on exercise (as preparation), or after a complex reading (as a reinforcement).
Prompts to try
Use these to draft your explainer scripts before entering Synthesia. ChatGPT or Claude can help structure the explanation.
Prompt 1 — Concept explainer script:
"Write a 2-minute explainer video script about [concept]. The audience is [your students]. Start with why this concept matters to them, then explain it step by step using a concrete example. Use simple language, short sentences, and a conversational tone. End with one key takeaway. Include [SLIDE] markers where a visual should appear."
Prompt 2 — Module introduction:
"Write a 90-second module introduction script for Module [X]: [Module Name] of my [course topic] course. Cover: what this module is about, what students will be able to do after completing it, and how it connects to the previous module. Warm, encouraging tone. No hype — just clear orientation."
Prompt 3 — Process walkthrough:
"Write a 2-minute video script walking through the process of [specific task]. The viewer is a beginner who has never done this before. Break it into 4-5 clear steps. For each step, describe what to do and what to look for to know you've done it correctly. Include [SLIDE] markers for screen captures or diagrams."
The human layer
Synthesia produces polished explainer videos efficiently. What it can't do is teach with presence. The kind of teaching where you lean forward, make eye contact, and say "I know this part is confusing — here's the thing most people miss." That spontaneous connection between teacher and student is what makes the difference between understanding a concept and truly internalizing it.
I'd recommend using AI explainers for the informational layer of your course — the parts where clarity and consistency matter most. Save your real presence for the moments that need real human connection: live sessions, recorded introductions where students meet you, feedback on student work, and the teaching moments where emotion and nuance carry the learning.
What it gets wrong
- Avatars can't demonstrate physical actions. If your course involves showing how to do something with your hands — drawing, cooking, instrument technique, physical therapy exercises — an AI avatar standing at a virtual podium won't work. You need real footage for anything physical. Synthesia is for explaining concepts, not demonstrating skills.
- Pacing is uniform. A good explainer naturally speeds up through familiar material and slows down for complex points. AI avatars deliver at a consistent pace regardless of content difficulty. You can partially compensate with script structure — shorter sentences for complex points create natural pauses — but the pacing never feels as responsive as a real instructor reading the room.
- The free tier is limited. Ten minutes per month means two or three short explainers. If you need a full set of module introductions (say, eight modules at two minutes each), you'll need the Starter plan at $29/month. Budget for at least one month of a paid plan during your initial course build, then downgrade if you don't need to produce regularly.
- Style consistency across videos requires intentional choices. If you generate explainers across different sessions, the avatar's positioning, background, and text overlay style can drift. Set up your preferred layout, avatar, and visual style as a Synthesia template and reuse it for every video to maintain consistency.
Frequently asked questions
How much does Synthesia cost for explainer videos?
Synthesia's free tier gives you 10 minutes of video per month — enough for one or two short explainers. The Starter plan ($29/month) adds more video minutes and better avatar options. The Creator plan ($89/month) includes custom avatars and priority rendering. For most course creators, the free tier or Starter plan covers supplementary explainer content. If you're producing multi-language versions of every module, the Creator plan becomes worthwhile.
What makes Synthesia different from other AI video tools?
Synthesia specializes in presenter-style videos with realistic AI avatars. While tools like Descript or Loom focus on editing real footage of you, Synthesia generates footage that doesn't require you to be on camera at all. The main advantage is speed and consistency — you can produce polished explainer videos from a script in minutes. The main limitation is that avatars lack the emotional range and spontaneity of a real presenter.
Can I add my own slides or screen recordings to Synthesia explainers?
Yes. Synthesia supports split-screen layouts where the avatar appears alongside slides, images, or screen recordings. This is the most effective format for explainer content — the avatar provides narration while the visual content shows what you're explaining. You can upload PowerPoint slides, images, or video clips to appear alongside the presenter.
Your explainer is ready — now build the lesson around it
You've got a polished explainer video that walks students through the concept. On Ruzuku, embed it directly in a lesson step alongside written instructions, practice exercises, and discussion prompts. The video explains; the rest of the lesson lets students apply what they've learned.
Related guides
- How to Create AI Avatar Course Videos Using Synthesia — broader guide to Synthesia for course content beyond explainers
- How to Record and Edit Course Videos Using Descript — record yourself and edit by editing text: the middle ground between AI and traditional video
- How to Create Voiceovers for Course Slides Using ElevenLabs — AI narration without the avatar, for slide-based explainers
- How to Create Your First Online Course — the complete guide from idea to launch