← Search

Yingjie Chen

11 accepted papers

2026

Sentiment-aware Rating-based Recommendation via Semantic-enhanced Item Alignment

IJCAI 2026

Leveraging review texts to mine deep user preferences is vital for recommendation. However, existing methods neglect the positive-negative counteraction and rely on noisy hard sentiment thresholds. Furthermore, the feature density asymmetry causes dense semantic features to overwhelm sparse collabor

Cited by 0Scholar
2025

Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation

ICCV 2025poster

Motion-controllable image animation is a fundamental task with a wide range of potential applications. Recent works have made progress in controlling camera or object motion via various motion representations, while they still struggle to support collaborative camera and object motion control with a…

Cited by 0SourcePDFScholar
2024

DYSON: Dynamic Feature Space Self-Organization for Online Task-Free Class Incremental Learning

CVPR 2024poster

In this paper we focus on a challenging Online Task-Free Class Incremental Learning (OTFCIL) problem. Different from the existing methods that continuously learn the feature space from data streams we propose a novel compute-and-align paradigm for the OTFCIL. It first computes an optimal geometry i.…

2024

Trend-Aware Supervision: On Learning Invariance for Semi-supervised Facial Action Unit Intensity Estimation

AAAI 2024technical

With the increasing need for facial behavior analysis, semi-supervised AU intensity estimation using only keyframe annotations has emerged as a practical and effective solution to relieve the burden of annotation. However, the lack of annotations makes the spurious correlation problem caused by AU c…

Cited by 0SourcePDFScholar
2022

Causal Intervention for Subject-Deconfounded Facial Action Unit Recognition

AAAI 2022technical

Subject-invariant facial action unit (AU) recognition remains challenging for the reason that the data distribution varies among subjects. In this paper, we propose a causal inference framework for subject-invariant facial action unit recognition. To illustrate the causal effect existing in AU recog…

Cited by 29SourcePDFScholar
2022

Improved Deep Unsupervised Hashing with Fine-grained Semantic Similarity Mining for Multi-Label Image Retrieval

IJCAI 2022poster

In this paper, we study deep unsupervised hashing, a critical problem for approximate nearest neighbor research. Most recent methods solve this problem by semantic similarity reconstruction for guiding hashing network learning or contrastive learning of hash codes. However, in multi-label scenarios,…

Cited by 16SourcePDFScholar
2022

On Mitigating Hard Clusters for Face Clustering

ECCV 2022poster

"Face clustering is a promising way to scale up face recognition systems using large-scale unlabeled face images. It remains challenging to identify small or sparse face image clusters that we call hard clusters, which is caused by the heterogeneity, i.e., high variations in size and sparsity, of th…

2022

Towards Unbiased Label Distribution Learning for Facial Pose Estimation Using Anisotropic Spherical Gaussian

ECCV 2022poster

"Facial pose estimation refers to the task of predicting face orientation from a single RGB image. It is an important research topic with a wide range of applications in computer vision. Label distribution learning (LDL) based methods have been recently proposed for facial pose estimation, which ach…

Cited by 33SourcePDFScholar
2021

Cross-Modal Representation Learning for Lightweight and Accurate Facial Action Unit Detection

RA-L 2021

In this letter, we focus on designing an effective method for lightweight and accurate facial action unit (AU) detection, which is essential for emotional communication in most human-robot interaction scenarios. AU detection is a delicate and challenging task because the subtle fleeting appearance c

Cited by 8SourceScholar
2021

DenserNet: Weakly Supervised Visual Localization Using Multi-Scale Feature Aggregation

AAAI 2021technical

In this work, we introduce a Denser Feature Network(DenserNet) for visual localization. Our work provides three principal contributions. First, we develop a convolutional neural network (CNN) architecture which aggregates feature maps at different semantic levels for image representations…

2021

SG-Net: Spatial Granularity Network for One-Stage Video Instance Segmentation

CVPR 2021poster

Video instance segmentation (VIS) is a new and critical task in computer vision. To date, top-performing VIS methods extend the two-stage Mask R-CNN by adding a tracking branch, leaving plenty of room for improvement. In contrast, we approach the VIS task from a new perspective and propose a one-sta…

Cited by 241PDFcodeScholar