OmniHuman·par ByteDanceAvatar parlant

OmniHuman 1.5

ByteDance OmniHuman 1.5: turns a portrait image and an audio track into a realistic, lip-synced talking video with facial expression and body motion. Billed per second of audio (≤60s).

Paramètres

ParamètreTypePar défautPlage ou options
seed
seed

Random seed; -1 for random.

number-1—
Resolution
output_resolution

Output video resolution (720 or 1080). Does not affect price.

select1080720, 1080

Entrées

Fichiers que ce modèle accepte en plus du prompt.

EntréeAccepteFichiers max.
Image
image_url
image1Obligatoire
Audio File
audio_url
audio1Obligatoire