ToonCrafter

Generates smooth in-between animation frames from two cartoon keyframes using video diffusion.

Open SourceSelf HostedOffline CapableGPU Required (24GB+ VRAM)
0.0 (0)

About

Traditional frame interpolation assumes small linear motions and breaks down on cartoons, where exaggerated movement and occlusion are the norm, so ToonCrafter treats inbetweening as a generative task: given two cartoon keyframes, a pretrained image-to-video diffusion prior adapted to the cartoon domain synthesizes up to 16 in-between frames at 512x320 resolution, taking around 24 seconds per clip on an A100 with 50 DDIM sampling steps. Optional sparse sketch guidance lets an animator steer the motion path with a few drawn strokes, and the same model supports reference-based sketch colorization. The work comes from CUHK and Tencent AI Lab and was published in the SIGGRAPH Asia 2024 journal track. Running it locally means a conda environment with Python 3.8.5 and roughly 24 GB of VRAM, which community optimizations bring closer to 10 GB, and ComfyUI nodes, Colab notebooks, and Windows ports exist. Code and weights are Apache-2.0 licensed with about 6,000 GitHub stars; the authors describe it as research exploration rather than a polished commercial product, without guaranteed results on every input.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Advanced (4/5)
License
Apache-2.0
Minimum VRAM
24 GB
Added
Jul 29, 2026

Related Tools

Animates a still human photo with 3D SMPL parametric motion guidance extracted from a driving video.

Open SourceSelf HostedOfflineGPU 20GB+
Advanced
0.0 (0)

Free markerless motion capture system that works with ordinary cameras and no special hardware.

Open SourceSelf HostedOffline
Easy
0.0 (0)

Audio-driven Tencent model that animates avatar images into emotion-controllable dialogue videos.

Open SourceSelf HostedOfflineGPU 10GB+
Advanced
0.0 (0)

ByteDance's audio-conditioned latent diffusion model for lip-syncing video to new speech.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

Audio-driven talking head animation from a single image.

Open SourceSelf HostedOfflineGPU 6GB+
Easy
0.0 (0)

Effective whole-body pose estimation with few-shot keypoint detection.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)
Browse all AI Animation & Motion tools