2023
MPE4G : Multimodal Pretrained Encoder for Co-Speech Gesture Generation
ICASSP 2023accepted
When virtual agents interact with humans, gestures are crucial to delivering their intentions with speech. Previous multimodal co-speech gesture generation models required encoded features of all modalities to generate gestures. If some input modalities are removed or contain noise, the model may no…