← Search

Zhenliang He

5 accepted papers

2026

Revisiting Visual Corruptions in LVLMs: A Shape-Texture Perspective on Model Failures

CVPR 2026

Large vision-language models (LVLMs) are highly vulnerable to visual corruptions, substantially compromising their reliability and limiting real-world deployment. Prior work has attributed this degradation primarily to insufficient visual grounding and overreliance on language priors. However, these

Cited by 0SourceScholar
2026

UniPercept: A Unified Diffusion Model for Generalizable Visual Perception

CVPR 2026

Diffusion models have shown impressive performance in generative tasks, demonstrating their ability to capture detailed structural and semantic information. Recently, these capabilities have been extended to visual understanding, with studies employing diffusion models as the backbone for various pe

Cited by 0SourcecodeScholar
2025

CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation

ICLR 2025poster

Recently, large-scale diffusion models have made impressive progress in text-to-image (T2I) generation. To further equip these T2I models with fine-grained spatial control, approaches like ControlNet introduce an extra network that learns to follow a condition image. However, for every single condit…

2019

S2GAN: Share Aging Factors Across Ages and Share Aging Trends Among Individuals

ICCV 2019oral

Generally, we human follow the roughly common aging trends, e.g., the wrinkles only tend to be more, longer or deeper. However, the aging process of each individual is more dominated by his/her personalized factors, including the invariant factors such as identity and mole, as well as the personaliz…

Cited by 57PDFScholar