← Search

Hongyang Du

7 accepted papers

2026

Distilling Geometry Priors for 3D-Consistent Video Generation

ICML 2026poster

While recent video diffusion models (VDMs) produce visually impressive results, they fundamentally struggle to maintain 3D structural consistency, often resulting in object deformation or spatial drift. We hypothesize that these failures arise because standard denoising objectives lack explicit ince…

Cited by 0SourceScholar
2026

Distilling Geometry Priors for 3D-Consistent Video Generation

ICML 2026poster

While recent video diffusion models (VDMs) produce visually impressive results, they fundamentally struggle to maintain 3D structural consistency, often resulting in object deformation or spatial drift. We hypothesize that these failures arise because standard denoising objectives lack explicit ince…

Cited by 0SourceScholar
2025

VideoHallu: Evaluating and Mitigating Multi-modal Hallucinations on Synthetic Video Understanding

NeurIPS 2025poster

Vision Language models (VLMs) have achieved remarkable success in video understanding tasks. Yet, a key question remains: Do they comprehend visual information or merely learn superficial mappings between visual and textual patterns? Understanding visual cues, particularly those related to physics…

Cited by 0SourcecodeScholar
2024

Generative Al-aided Joint Training-free Secure Semantic Communications via Multi-modal Prompts

ICASSP 2024accepted

Semantic communication (SemCom) holds promise for reducing network resource consumption while achieving the communications goal. However, the computational overheads in jointly training semantic encoders and decoders—and the subsequent deployment in network devices—are overlooked. Recent advances in…

Cited by 0SourceScholar
2024

Scalable Federated Unlearning via Isolated and Coded Sharding

IJCAI 2024poster

Federated unlearning has emerged as a promising paradigm to erase the client-level data effect without affecting the performance of collaborative learning models. However, the federated unlearning process often introduces extensive storage overhead and consumes substantial computational resources, t…

2023

Multi-Agent Reinforcement Learning for Covert Semantic Communications over Wireless Networks

ICASSP 2023accepted

In this paper, a covert semantic communication framework is proposed for image transmission over wireless networks. In the proposed framework, devices extract and selectively transmit semantic information of image data to a base station (BS). The semantic information consists of the objects in the i…

Cited by 0SourceScholar