Avatar Video

Animates a person from a single image so that they lip-sync a provided audio track. The process is asynchronous and will require an additional call to get the result. See below for more details.

Recent Requests
Log in to see full request history
TimeStatusUser Agent
Retrieving recent requests…
LoadingLoading…
Body Params
uri
required
length between 1 and 2083

Source image URL of the person to animate.

uri
required
length between 1 and 2083

URL of the audio track the avatar lip-syncs to. The length of the track drives the duration of the generated video, and with it the price.
Depending on the selected model, WAV and MP3 may be the only accepted formats, and the track may be capped at 60 seconds for 720p or 30 seconds for 1080p.

string
length between 1 and 10000

Text prompt guiding motion, emotion and camera movement.

string
enum

Video output quality/resolution. Mapped to the closest tier the underlying model supports. Options are as follows:

  • 480p
  • 720p
  • 1080p
  • 4k
Allowed:
string | null
enum
Defaults to urn:air:heygen:model:heygen:avatar-iv@1

Optionally choose a specific AI model to use for this video generation.
If not specified, the model configured for your API key is used, and where none is configured a default model will be applied. Please note that the default model may change over time as Picsart continues to improve performance and accuracy. Any change to the default will be made only after thorough testing and validation to ensure it delivers better results.
If you require consistent behavior or wish to evaluate different models on your own, we recommend explicitly setting this parameter.

Allowed:
Responses

Language
Credentials
Header
LoadingLoading…
Response
Click Try It! to start a request and see the response here! Or choose an example:
application/json