2021
Learning Pose-Adaptive Lip Sync with Cascaded Temporal Convolutional Network
ICASSP 2021accepted
Speech-driven lip sync has become a promising technique for generating and editing talking-head videos. These studies mainly use 3D morphable models or 2D facial landmarks as the intermediate face representations. However, 2D-based methods have been stagnant recently due to their inability to handle…