← Search

Ling.Yu Zhu

8 accepted papers

2026

Beyond Cosine Similarity: Magnitude-Aware CLIP for No-Reference Image Quality Assessment

AAAI 2026technical

Recent efforts have repurposed the Contrastive Language-Image Pre-training (CLIP) model for No-Reference Image Quality Assessment (NR-IQA) by measuring the cosine similarity between the image embedding and textual prompts such as "a good photo" or "a bad photo." However, this semantic similarity ove

Cited by 0SourcePDFScholar
2026

CADC: Content Adaptive Diffusion-Based Generative Image Compression

CVPR 2026

Diffusion-based generative image compression has demonstrated remarkable potential for achieving realistic reconstruction at ultra-low bitrates. The key to unlocking this potential lies in making the entire compression process content-adaptive, ensuring that the encoder's representation and the deco

Cited by 0SourceScholar
2026

Mitigating Perception Bias: A Training-Free Approach to Enhance LMM for Image Quality Assessment

AAAI 2026technical

Despite the impressive performance of large multimodal models (LMMs) in high-level visual tasks, their capacity for image quality assessment (IQA) remains limited. One main reason is that LMMs are primarily trained for high-level tasks (e.g., image captioning), emphasizing unified image semantics ex

Cited by 0SourcePDFScholar
2025

Boosting Multi-View Indoor 3D Object Detection via Adaptive 3D Volume Construction

ICCV 2025poster

This work presents SGCDet, a novel multi-view indoor 3D object detection framework based on adaptive 3D volume construction. Unlike previous approaches that restrict the receptive field of voxels to fixed locations on images, we introduce a geometry and context aware aggregation module to integrate…

2025

Noise2Score3D: Tweedie's Approach for Unsupervised Point Cloud Denoising

ICCV 2025poster

Building on recent advances in Bayesian statistics and image denoising, we propose Noise2Score3D, a fully unsupervised framework for point cloud denoising. Noise2Score3D learns the score function of the underlying point cloud distribution directly from noisy data, eliminating the need for clean data…

Cited by 0SourcePDFScholar
2024

Adaptive Image Quality Assessment via Teaching Large Multimodal Model to Compare

NeurIPS 2024spotlight

While recent advancements in large multimodal models (LMMs) have significantly improved their abilities in image quality assessment (IQA) relying on absolute quality rating, how to transfer reliable relative quality comparison outputs to continuous perceptual quality scores remains largely unexplore…

2024

Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video Enhancement

ECCV 2024poster

"Obtaining pairs of low/normal-light videos, with motions, is more challenging than still images, which raises technical issues and poses the technical route of unpaired learning as a critical role. This paper makes endeavors in the direction of learning for low-light video enhancement without using…

2022

Multi-Party Empathetic Dialogue Generation: A New Task for Dialog Systems

ACL 2022long

Empathetic dialogue assembles emotion understanding, feeling projection, and appropriate response generation. Existing work for empathetic dialogue generation concentrates on the two-party conversation scenario. Multi-party dialogues, however, are pervasive in reality. Furthermore, emotion and sensi…

Cited by 17SourcePDFScholar