Champ

Animates a still human photo with 3D SMPL parametric motion guidance extracted from a driving video.

Open SourceSelf HostedOffline CapableGPU Required (20GB+ VRAM)
0.0 (0)

About

Champ animates a person in a single reference image by extracting SMPL parametric body sequences from a driving video and rendering depth maps, surface normals, semantic segmentation, and DWPose skeletons as conditioning signals for a latent diffusion generator. Because the guidance comes from a unified 3D body model rather than 2D keypoints alone, body shape and motion stay consistent even under large pose changes, which was the core contribution of the ECCV 2024 paper. The project comes from Fudan University's generative vision lab, the same group behind the Hallo talking-head models, and ships pretrained weights on Hugging Face along with full two-stage training code and sample training data. Running inference requires Python 3.10, CUDA 12.1, and roughly 20 GB of VRAM for a 250-frame sequence on an A100 or RTX 3090, with a frame-range option to trim memory use on smaller cards. Community wrappers exist for ComfyUI and Blender. Code and weights are MIT licensed, so commercial use is permitted, and the repository holds about 4,300 GitHub stars.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Advanced (4/5)
License
MIT
Minimum VRAM
20 GB
Added
Jul 29, 2026

Related Tools

Free markerless motion capture system that works with ordinary cameras and no special hardware.

Open SourceSelf HostedOffline
Easy
0.0 (0)

Audio-driven Tencent model that animates avatar images into emotion-controllable dialogue videos.

Open SourceSelf HostedOfflineGPU 10GB+
Advanced
0.0 (0)

ByteDance's audio-conditioned latent diffusion model for lip-syncing video to new speech.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

Audio-driven talking head animation from a single image.

Open SourceSelf HostedOfflineGPU 6GB+
Easy
0.0 (0)

Real-time high-quality lip-sync model for audio-driven talking face generation.

Open SourceSelf HostedOfflineGPU 6GB+
Intermediate
0.0 (0)

Effective whole-body pose estimation with few-shot keypoint detection.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)
Browse all AI Animation & Motion tools