LatentSync1.6, an end-to-end lip-sync method

3 pointsposted 10 hours ago
by BruceWok

1 Comments

BruceWok

10 hours ago

LatentSync is a cutting-edge, open-source lip-synchronization framework powered by Audio-Conditioned Latent Diffusion Models. By integrating Whisper audio embeddings with advanced temporal alignment (TREPA), it transforms arbitrary audio and video inputs into photorealistic, high-resolution (512x512) talking head videos.