← Search

Peilin Chen

7 accepted papers

2026

Dual Graph Regularized Deep Unfolding Network for Guided Depth Map Super-resolution

CVPR 2026

Depth map super-resolution with color guidance is a fundamental task in computer vision that aims to reconstruct high-resolution depth maps by leveraging structural correlations from corresponding guidance images. Recently, with the development of deep learning techniques, the performance of guided

Cited by 0SourceScholar
2026

When Privacy Meets Recovery: The Overlooked Half of Surrogate-Driven Privacy Preservation for MLLM Editing

AAAI 2026technical

Privacy leakage in Multimodal Large Language Models (MLLMs) has long been an intractable problem. Existing studies, though effectively obscure private information in MLLMs, often overlook the evaluation of authenticity and recovery quality of user privacy. To this end, this work uniquely focuses on

Cited by 0SourcePDFScholar
2025

An Information-Theoretic Regularizer for Lossy Neural Image Compression

ICCV 2025poster

Lossy image compression networks aim to minimize the latent entropy of images while adhering to specific distortion constraints. However, optimizing the neural network can be challenging due to its nature of learning quantized latent representations. In this paper, our key finding is that minimizing…

Cited by 0SourcePDFScholar
2025

CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models

ICML 2025poster

Recent advances in clinical AI have enabled remarkable progress across many clinical domains. However, existing benchmarks and models are primarily limited to a small set of modalities and tasks, which hinders the development of large-scale multimodal methods that can make holistic assessments of pa…

2025

Making Old Film Great Again: Degradation-aware State Space Model for Old Film Restoration

CVPR 2025poster

Unlike modern native digital videos, the restoration of old films requires addressing specific degradations inherent to analog sources. However, existing specialized methods still fall short compared to general video restoration techniques. In this work, we propose a new baseline to re-examine the c…

2025

QoQ-Med: Building Multimodal Clinical Foundation Models with Domain-Aware GRPO Training

NeurIPS 2025oral

Clinical decision‑making routinely demands reasoning over heterogeneous data, yet existing multimodal language models (MLLMs) remain largely vision‑centric and fail to generalize across clinical specialties. To bridge this gap, we introduce QoQ-Med-7B/32B, the first open generalist clinical foundati…

Cited by 0SourceScholar
2022

SMINet: State-Aware Multi-Aspect Interests Representation Network for Cold-Start Users Recommendation

AAAI 2022technical

Online travel platforms (OTPs), e.g., bookings.com and Ctrip.com, deliver travel experiences to online users by providing travel-related products. Although much progress has been made, the state-of-the-arts for cold-start problems are largely sub-optimal for user representation, since they do not ta…