LatentSync
Taming Stable Diffusion for Lip Sync!
About this project
LatentSync 🔥 Updates — 2025/06/11: We released LatentSync 1.6, which is trained on 512 $\times$ 512 resolution videos to mitigate the blurriness problem. Watch the demo here. — 2025/03/14: We released LatentSync 1.5, which (1) improves temporal consistency via adding temporal layer, (2) improves performance on Chinese videos and (3) reduces the VRAM requirement of the stage2 training to 20 GB through a series of optimizations. Learn more details here. 📖 Introduction We present LatentSync, an end-to-end lip-sync method based on audio-conditioned latent diffusion models without any intermediate motion representation, diverging from previous diffusion-based lip-sync…
Technologies
Project health
GitHub
Reviews
Built by
Maintain bytedance/LatentSync? Claiming verifies admin access through your GitHub account and gives you control of this listing.
