← Back to all generators

cjwbw/sadtalker

Stylized Audio-Driven Single Image Talking Face Animation

Capabilities

No capability data available

Cost

Community model (estimated from hardware time)

Input Parameters

driven_audiorequiredstring

Upload the driven audio, accepts .wav and .mp4 file

source_imagerequiredstring

Upload the source image, it can be video.mp4 or picture.png

expression_scalenumber

a larger value will make the expression motion stronger

Default: 1
facerenderstring

Choose face render

Default: "facevid2vid"
facevid2vidpirender
pose_styleinteger

Pose style

Default: 0min: 0, max: 45
preprocessstring

Choose how to preprocess the images

Default: "crop"
cropresizefullextcropextfull
size_of_imageinteger

Face model resolution

Default: 256
256512
still_modeboolean

Still Mode (fewer head motion, works with preprocess 'full')

Default: true
use_enhancerboolean

Use GFPGAN as Face enhancer

Default: false
use_eyeblinkboolean

Use eye blink

Default: true
Version: a519cc0cfebaUpdated: 7/25/2026159.4K runs