Hallo
Hierarchical audio-driven visual synthesis for portrait animation.
About
Hallo, from Fudan University and collaborators, generates portrait animation videos driven by audio using a hierarchical audio-driven visual synthesis approach. From a reference portrait and a speech clip it produces synchronized lip movements, facial expressions, and head motion. The repository provides inference scripts and pretrained weights, and the community has built additional resources around it. Released as an open-source research project.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- AI Animation & Motion
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Intermediate (3/5)
- Minimum VRAM
- 8 GB
- Added
- Apr 3, 2026
Related Tools
Animates a still human photo with 3D SMPL parametric motion guidance extracted from a driving video.
Free markerless motion capture system that works with ordinary cameras and no special hardware.
Audio-driven Tencent model that animates avatar images into emotion-controllable dialogue videos.
ByteDance's audio-conditioned latent diffusion model for lip-syncing video to new speech.
Audio-driven talking head animation from a single image.
Effective whole-body pose estimation with few-shot keypoint detection.