SadTalker

SadTalker

Audio-driven talking face from one image

Description

Making a portrait photo talk usually yields mismatched lips or a frozen head. SadTalker learns realistic 3D motion coefficients from audio to animate a single image into a natural talking-head video.

Lips, expressions and head pose follow the voice across styles, with a web UI and a Stable Diffusion WebUI extension; CVPR 2023.

Features



Single image:One photo.

Lip sync:Audio-driven.

Head motion:3D coefficients.

WebUI:SD extension.