All courses

AI Video: Avatars, Text-to-Video & Lip Sync

The three kinds of AI video and which you actually need, avatar work that does not feel cheap, generative footage limits, and localisation done in the right order.

Most disappointment with AI video comes from reaching for generative footage when the job was an avatar reading a script. This course separates the three technologies (avatar and presenter, generative, and assisted editing, which is the least glamorous and highest return), then teaches each properly. You will learn what makes avatar video watchable, including cutting away every 5 to 10 seconds and why EY India found 64% of Indian viewers accept AI avatars when content is useful and clearly labelled. You will prompt generative clips as commissioned stock footage with realistic keep rates, caption to real subtitle specifications, and localise in the correct order: translate, human check, then render.

What you'll learn

Course content

  1. 1. The three kinds of AI video, and which you need (15 min)
  2. 2. The 2026 tool landscape, and why it will change again (15 min)
  3. 3. Avatar video that does not feel cheap (15 min)
  4. 4. Generative footage: prompting and its limits (15 min)
  5. 5. Assembly, subtitles, and localisation (15 min)
  6. 6. A production workflow you can repeat (15 min)

Related courses