SkyReels-V3

Skywork's open-weight video models for reference-to-video, footage extension, and audio-driven avatars.

Open SourceSelf HostedOffline CapableGPU Required (24GB+ VRAM)
0.0 (0)

About

Skywork's third SkyReels generation splits video work across three open-weight pipelines: a 14B reference-to-video model that preserves the identity of 1 to 4 people or objects from supplied images, a 14B video extension model that continues existing footage with cinematographic transitions, and a 19B audio-driven talking avatar model, all generating at 720p. Inference code landed on GitHub in January 2026 with checkpoints auto-downloaded from Hugging Face or ModelScope; the stack wants Python 3.12 and CUDA 12.8, and cards under 24 GB can still run it by enabling FP8 weight-only quantization with block offload or dropping output to 540p or 480p. The weights ship under the Skywork Community License, which permits commercial use, and the same models back Skywork's paid API for teams that skip local hosting. Multi-subject consistency is the headline capability: reference conditioning makes this one of the few open options for keeping specific faces and products stable across generated shots, a task where prompt-only video models routinely drift.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Freemium
Platform
Hybrid
Difficulty
Advanced (4/5)
License
Skywork Community License
Minimum VRAM
24 GB
Added
Aug 24, 2026

Related Tools

Featured

Open-source video generation model by Tencent with text and image conditioning.

Open SourceSelf HostedOfflineGPU 24GB+
Advanced

Image-to-video generation model by Alibaba DAMO Academy.

Open SourceSelf HostedOfflineGPU 12GB+
Advanced

Updated CogVideo model by Zhipu AI with improved video quality.

Open SourceSelf HostedOfflineGPU 16GB+
Advanced

Infinite-length music-driven video generation with visual conditioning.

Open SourceSelf HostedOfflineGPU 12GB+
Advanced

Text-to-video generation framework with cascaded latent diffusion.

Open SourceSelf HostedOfflineGPU 16GB+
Advanced

Open-source text-to-video model by Zhipu AI/Tsinghua with 2B and 5B variants.

Open SourceSelf HostedOfflineGPU 12GB+
Advanced
Browse all Video Generation tools