SwapTalk: Audio-Driven Talking Face Generation with One-Shot Customization in Latent Space
Combining face-swapping with lip synchronization offers a cost-effective solution for generating customized talking faces. However, directly cascading existing models can introduce significant interference and reduce video clarity due to limited interaction space in the low-level RGB domain. To solv…