← Search

Yunan Li

10 accepted papers

2025

Enhanced Contrastive Learning with Multi-view Longitudinal Data for Chest X-ray Report Generation

CVPR 2025poster

Automated radiology report generation offers an effective solution to alleviate radiologists' workload. However, most existing methods focus primarily on single or fixed-view images to model current disease conditions, which limits diagnostic accuracy and overlooks disease progression. Although some…

2025

Gaussian-Face: Talking Head Generation with Hybrid Density via 3D Gaussian Splatting

ICASSP 2025accepted

In recent years, audio-driven neural radiance field (NeRF)-based talking head generation techniques have achieved impressive results. However, these methods still have some limitations, such as unsynchronized lip movements and visual jitter. Recently, 3D Gaussian splatting has gradually replaced NeR…

Cited by 0SourceScholar
2025

KAN-Face: Efficient Resource Usage and Precision Lip-Sync in Talking Head Generation

ICASSP 2025accepted

Despite significant progress in NeRF-based talking head generation, problems like poor lip synchronization and inefficient resource usage remain. To solve these, we propose KANFace, a lightweight framework. In preprocessing, we introduce a Lip-Sync Enhancement Module that uses Wav2Lip to extract hig…

Cited by 0SourceScholar
2025

What is Stigma Attributed to? A Theory-Grounded, Expert-Annotated Interview Corpus for Demystifying Mental-Health Stigma

ACL 2025long

Mental-health stigma remains a pervasive social problem that hampers treatment-seeking and recovery. Existing resources for training neural models to finely classify such stigma are limited, relying primarily on social-media or synthetic data without theoretical underpinnings. To remedy this gap, we…

2023

Learning Robust Representations with Information Bottleneck and Memory Network for RGB-D-based Gesture Recognition

ICCV 2023poster

Although previous RGB-D-based gesture recognition methods have shown promising performance, researchers often overlook the interference of task-irrelevant cues like illumination and background. These unnecessary factors are learned together with the predictive ones by the network and hinder accurate…

Cited by 8PDFScholar
2022

Fidelity Evaluation of Virtual Traffic Based on Anomalous Trajectory Detection

IROS 2022poster

Measuring the fidelity of synthesized virtual traffic has become an important and fundamental concern for evaluating the performance of different traffic simulation techniques and applications of autonomous vehicle testing. In this work, we propose a novel method to evaluate the fidelity of any traj…

Cited by 1SourceScholar
2021

Regional Attention with Architecture-Rebuilt 3D Network for RGB-D Gesture Recognition

AAAI 2021technical

Human gesture recognition has drawn much attention in the area of computer vision. However, the performance of gesture recognition is always influenced by some gesture-irrelevant factors like the background and the clothes of performers. Therefore, focusing on the regions of hand/arm is important to…

2019

LAP-Net: Level-Aware Progressive Network for Image Dehazing

ICCV 2019poster

In this paper, we propose a level-aware progressive network (LAP-Net) for single image dehazing. Unlike previous multi-stage algorithms that generally learn in a coarse-to-fine fashion, each stage of LAP-Net learns different levels of haze with different supervision. Then the network can progressive…

Cited by 86PDFScholar