← Search

Vandana Rajan

2 accepted papers

2022

Is Cross-Attention Preferable to Self-Attention for Multi-Modal Emotion Recognition?

ICASSP 2022accepted

Humans express their emotions via facial expressions, voice intonation and word choices. To infer the nature of the underlying emotion, recognition models may use a single modality, such as vision, audio, and text, or a combination of modalities. Generally, models that fuse complementary information…

Cited by 0SourceScholar
2021

Robust Latent Representations Via Cross-Modal Translation and Alignment

ICASSP 2021accepted

Multi-modal learning relates information across observation modalities of the same physical phenomenon to leverage complementary information. Most multi-modal machine learning methods require that all the modalities used for training are also available for testing. This is a limitation when signals…

Cited by 0SourceScholar