ToonCrafter
Generates smooth in-between animation frames from two cartoon keyframes using video diffusion.
About
Traditional frame interpolation assumes small linear motions and breaks down on cartoons, where exaggerated movement and occlusion are the norm, so ToonCrafter treats inbetweening as a generative task: given two cartoon keyframes, a pretrained image-to-video diffusion prior adapted to the cartoon domain synthesizes up to 16 in-between frames at 512x320 resolution, taking around 24 seconds per clip on an A100 with 50 DDIM sampling steps. Optional sparse sketch guidance lets an animator steer the motion path with a few drawn strokes, and the same model supports reference-based sketch colorization. The work comes from CUHK and Tencent AI Lab and was published in the SIGGRAPH Asia 2024 journal track. Running it locally means a conda environment with Python 3.8.5 and roughly 24 GB of VRAM, which community optimizations bring closer to 10 GB, and ComfyUI nodes, Colab notebooks, and Windows ports exist. Code and weights are Apache-2.0 licensed with about 6,000 GitHub stars; the authors describe it as research exploration rather than a polished commercial product, without guaranteed results on every input.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- AI Animation & Motion
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Advanced (4/5)
- License
- Apache-2.0
- Minimum VRAM
- 24 GB
- Added
- Jul 29, 2026
Related Tools
Animates a still human photo with 3D SMPL parametric motion guidance extracted from a driving video.
Free markerless motion capture system that works with ordinary cameras and no special hardware.
Audio-driven Tencent model that animates avatar images into emotion-controllable dialogue videos.
ByteDance's audio-conditioned latent diffusion model for lip-syncing video to new speech.
Audio-driven talking head animation from a single image.
Effective whole-body pose estimation with few-shot keypoint detection.