Creating a Single Talking Actor

The Talking Actor tool generates a lip-synced video by animating an actor from a reference image or video using a voice script via Text-to-Speech (TTS) with voice cloning or an uploaded voice.

Required Elements

  • Image / Video: Reference image or video of the actor. Supported formats: PNG, JPG, JPEG, WebP, MP4, WMV, MOV, and AVI.
  • Voice: Load an audio file for the actor’s voice, or use a preset or reference audio for TTS voice generation. Supported formats: MP3 and WAV.

AI Settings

Select the Talking Actor Tool from the left AI Toolbar and complete the tasks in each section under the Single Talking Actor tab. You can collapse or expand a section by clicking its caption.

Ensure both an image / video and voice audio are provided to enable the GENERATE button. Note the task price displayed above the button before clicking.

Track progress on the AI Render View or via the new entry in History on the right side of the AI Workspace. When the submission completes successfully, the AI-generated video will appear. Click Play to play the video, Loop to repeat it, or drag the playhead to a specific point.

Next, use the Video options in the Quick Access Menu at the top of the AI Render View to edit the video or upscale the generated output to a higher resolution.