← Search

Yifan He

10 accepted papers

2026

D^3FER: Dual Channel and Dual Branch Network for Robust Facial Expression Recognition under Dual Challenges

CVPR 2026

Facial expression recognition (FER) in the wild is challenged by co-occurring visual perturbations (e.g., occlusions, pose variations) and label noise. Existing methods often address these issues in isolation, failing to handle their compound effects effectively. To this end, we propose D^3FER (Dual

Cited by 0SourcecodeScholar
2025

Exploring Inter-Variate and Long-Term Dependencies to Boost Multivariate Time Series Forecasting

ICASSP 2025accepted

Multivariate Time Series Forecasting (MTSF) is a critical task in various domains, and Large Language Models (LLMs) for MTSF have recently received considerable attention. Despite significant progress in large-scale time series models, particularly in fine-tuning pre-trained LLMs for MTSF, there are…

Cited by 0SourceScholar
2024

Diversity-Authenticity Co-constrained Stylization for Federated Domain Generalization in Person Re-identification

AAAI 2024technical

This paper tackles the problem of federated domain generalization in person re-identification (FedDG re-ID), aiming to learn a model generalizable to unseen domains with decentralized source domains. Previous methods mainly focus on preventing local overfitting. However, the direction of diversifyin…

2024

Multivariate Time Series Forecasting with Causal-Temporal Attention Network

ICASSP 2024accepted

The task of multivariate time series (MTS) forecasting has attracted much attention in recent years. However, most existing methods overlook the causal relationship among different variables, which may lead to inaccurate forecasting results. In this paper, we incorporate causality into the forecasti…

Cited by 0SourceScholar
2024

Selective Domain-Invariant Feature for Generalizable Deepfake Detection

ICASSP 2024accepted

With diverse presentation forgery methods emerging continually, detecting the authenticity of images has drawn growing attention. Although existing methods have achieved impressive accuracy in training dataset detection, they still perform poorly in the unseen domain and suffer from forgery of irrel…

Cited by 0SourceScholar
2024

Voice Toxicity Detection Using Multi-Task Learning

ICASSP 2024accepted

Social communication systems must identify toxic voice audio to support moderation that protects the safety and civility of their communities. Toxicity classification for voice depends on both audio style, such as volume and tone, and content, such as the words in the speech individually and in cont…

Cited by 0SourceScholar
2023

Transplayer: Timbre Style Transfer with Flexible Timbre Control

ICASSP 2023accepted

Music timbre style transfer aims at replacing the instrument timbre in a solo recording with another instrument, while preserving the musical content. Existing GAN-based methods can only achieve timbre style transfer between two given timbres. Inspired by the practice in voice conversion, we propose…

Cited by 0SourceScholar
2022

An Efficient Framework for Detection and Recognition of Numerical Traffic Signs

ICASSP 2022accepted

Due to the variety of categories and uneven distribution of available samples, automatic traffic sign detection and recognition is still a challenging task. For those categories with less training data, existing deep learning methods cannot achieve desirable performance, and the overall detection ef…

Cited by 0SourceScholar
2020

Deep Interleaved Network for Single Image Super-Resolution with Asymmetric Co-Attention

IJCAI 2020poster

Recently, Convolutional Neural Networks (CNN) based image super-resolution (SR) have shown significant success in the literature. However, these methods are implemented as single-path stream to enrich feature maps from the input for the final prediction, which fail to fully incorporate former low-le…

Cited by 0SourcePDFScholar