Featured Tool

SadTalker

Audio-driven talking head animation from a single image.

Open SourceSelf HostedOffline CapableGPU Required (6GB+ VRAM)
0.0 (0)

About

SadTalker generates a talking-head video from a single portrait image and an audio clip by predicting 3D motion coefficients that drive natural head movement and lip sync. Decoupling expression and pose from the audio improves realism over earlier single-image methods. The project provides a Colab notebook, a Hugging Face Space, and a Stable Diffusion WebUI extension. Developed at Xi'an Jiaotong University and released under the MIT license.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Easy (2/5)
License
MIT
Minimum VRAM
6 GB
Added
Apr 3, 2026

Related Tools

Animates a still human photo with 3D SMPL parametric motion guidance extracted from a driving video.

Open SourceSelf HostedOfflineGPU 20GB+
Advanced
0.0 (0)

Free markerless motion capture system that works with ordinary cameras and no special hardware.

Open SourceSelf HostedOffline
Easy
0.0 (0)

Audio-driven Tencent model that animates avatar images into emotion-controllable dialogue videos.

Open SourceSelf HostedOfflineGPU 10GB+
Advanced
0.0 (0)

ByteDance's audio-conditioned latent diffusion model for lip-syncing video to new speech.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)

Real-time high-quality lip-sync model for audio-driven talking face generation.

Open SourceSelf HostedOfflineGPU 6GB+
Intermediate
0.0 (0)

Effective whole-body pose estimation with few-shot keypoint detection.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)
Browse all AI Animation & Motion tools