← Search

Liejun Wang

9 accepted papers

2026

All in One: Unifying Deepfake Detection, Tampering Localization, and Source Tracing with a Robust Landmark-Identity Watermark

CVPR 2026

With the rapid advancement of deepfake technology, malicious face manipulations pose a significant threat to personal privacy and social security. However, existing proactive forensics methods typically treat deepfake detection, tampering localization, and source tracing as independent tasks, lackin

Cited by 0SourcecodeScholar
2026

The Double Dilemma in Multi-Task Radiology Report Generation: A Gradient Dynamics Analysis and Solution

ICML 2026poster

While multi-task learning based automatic radiology report generation (RRG) is widely adopted to ensure clinical consistency, most focus on architectural designs yet remain limited to coarse linear scalarization strategies. These strategies can not effectively balance the hard constraints of discrim…

Cited by 0SourceScholar
2025

Hierarchical Perceptual Distillation Network for Lightweight Image Super-Resolution Reconstruction

ICASSP 2025accepted

Recently, the image super-resolution (SR) has made remarkable progress. However, due to the proliferation of resource-constrained scenarios, the computational-intensive SR technology is limited in portable devices. Therefore, high efficiency and lightweight become the key factors of image SR in the…

Cited by 0SourceScholar
2025

Modality-Invariant Bidirectional Temporal Representation Distillation Network for Missing Multimodal Sentiment Analysis

ICASSP 2025accepted

Multimodal Sentiment Analysis (MSA) integrates diverse modalities—text, audio, and video—to comprehensively analyze and understand individuals’ emotional states. However, the real-world prevalence of incomplete data poses significant challenges to MSA, mainly due to the randomness of modality missin…

Cited by 0SourceScholar
2025

Similarity Memory Prior is All You Need for Medical Image Segmentation

ICCV 2025poster

In recent years, it has been found that "grandmother cells" in the primary visual cortex (V1) of macaques can directly recognize visual input with complex shapes. This inspires us to examine the value of these cells in promoting the research of medical image segmentation. In this paper, we design a…

2023

Scoreformer: Score Fusion-Based Transformers for Weakly-Supervised Violence Detection

ICASSP 2023accepted

Violence detection is an application of anomaly detection, which is used to detect violence content in video clips. Using multimodal as input can improve the performance of violence detection. However, the existing MML Transformers-based fusion methods do not take into account the differences betwee…

Cited by 0SourceScholar
2021

Bidirectional Focused Semantic Alignment Attention Network for Cross-Modal Retrieval

ICASSP 2021accepted

Cross-modal retrieval is a very challenging and significant task in intelligent understanding. Researchers have tried to capture modal semantic information through a weighted attention mechanism. Still, they cannot eliminate irrelevant semantic information's negative effects and cannot capture fine-…

Cited by 0SourceScholar