Skip to main content
Animate a still image of a person into a lip-synced talking video using HeyGen’s Avatar IV technology. You can drive the animation with a text script that HeyGen converts to speech, or provide your own audio for the avatar to lip-sync.

Inputs

Common Inputs

Script Inputs

Shown when speech is "script".

Audio Inputs

Shown when speech is "audio". Note: When speech is "script", text must be specified, and a voice is required via the voice selector (choosing anything other than the avatar’s default voice) or a custom_voice_id. When speech is "audio", audio is required instead.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 2181066a8c6191cfcaa15ece4f89a16c37e76aa22763d6df4007baa20336f05a