← Search

Yinfeng Yu

4 accepted papers

2025

Modality-Invariant Bidirectional Temporal Representation Distillation Network for Missing Multimodal Sentiment Analysis

ICASSP 2025accepted

Multimodal Sentiment Analysis (MSA) integrates diverse modalities—text, audio, and video—to comprehensively analyze and understand individuals’ emotional states. However, the real-world prevalence of incomplete data poses significant challenges to MSA, mainly due to the randomness of modality missin…

Cited by 0SourceScholar
2023

Measuring Acoustics with Collaborative Multiple Agents

IJCAI 2023poster

As humans, we hear sound every second of our life. The sound we hear is often affected by the acoustics of the environment surrounding us. For example, a spacious hall leads to more reverberation. Room Impulse Responses (RIR) are commonly used to characterize environment acoustics as a function of t…

Cited by 3SourcePDFScholar
2023

SRTNET: Time Domain Speech Enhancement via Stochastic Refinement

ICASSP 2023accepted

Diffusion model, as a new generative model which is very popular in image generation and audio synthesis, is rarely used in speech enhancement. In this paper, we use the diffusion model as a module for stochastic refinement. We propose SRTNet, a novel method for speech enhancement via Stochastic Ref…

Cited by 0SourceScholar
2022

Sound Adversarial Audio-Visual Navigation

ICLR 2022poster

Audio-visual navigation task requires an agent to find a sound source in a realistic, unmapped 3D environment by utilizing egocentric audio-visual observations. Existing audio-visual navigation works assume a clean environment that solely contains the target sound, which, however, would not be suita…