← Search

Shi Chen

19 accepted papers

2026

InterLight: Leveraging Intrinsic Illumination Priors for Low-Light Image Enhancement

IJCAI 2026

Low-Light Image Enhancement (LLIE) has long been a challenging problem in low-level vision, as insufficient illumination often leads to low contrast, detail loss, and noise. Recent studies show that deep learning-based Retinex theory can effectively decouple illumination and reflectance. However, ex

Cited by 0Scholar
2026

Preference-Enhanced Reinforcement Learning for Pluralistic Image Inpainting

ICML 2026poster

Existing image inpainting frameworks rely on strictly supervised training paradigms, often suffering from an over-reliance on ground-truth reconstruction, which leads to conservative outputs with misaligned creativity and limited diversity. To this end, we propose the first framework to explore Grou…

Cited by 0SourceScholar
2026

UI-Lens: Assessing General MLLMs' Potential to Automate UI Display Quality Assurance

CVPR 2026

User Interface (UI) display defect detection poses challenges far beyond UI understanding, requiring fine-grained element boundary understanding, missing-content detection, and reasoning about sequential interface semantic consistency. However, the capabilities of multimodal large language models (M

Cited by 0SourceScholar
2025

HSRMamba: Contextual Spatial-Spectral State Space Model for Single Hyperspectral Image Super-Resolution

IJCAI 2025

Mamba has demonstrated exceptional performance in visual tasks due to its powerful global modeling capabilities and linear computational complexity, offering considerable potential in hyperspectral image super-resolution (HSISR). However, in HSISR, Mamba faces challenges as transforming images into

2025

ProDyG: Progressive Dynamic Scene Reconstruction via Gaussian Splatting from Monocular Videos

NeurIPS 2025poster

Achieving truly practical dynamic 3D reconstruction requires online operation, global pose and map consistency, detailed appearance modeling, and the flexibility to handle both RGB and RGB-D inputs. However, existing SLAM methods typically merely remove the dynamic parts or require RGB-D input, whil…

Cited by 0SourceScholar
2024

A Novel Geometrical Structure Robot Hand for Linear-parallel Pinching and Coupled Self-adaptive Hybrid Grasping

IROS 2024poster

Current robot hand grippers capable of self-adaptive or coupled grasping often cannot perform linear-parallel pinching at the physical end of the gripper, which is widely used in industrial applications. For this reason, this paper introduces a gripper with hybrid grasping modes— the LPCSA hand. It…

Cited by 1SourceScholar
2023

Divide and Conquer: Answering Questions With Object Factorization and Compositional Reasoning

CVPR 2023poster

Humans have the innate capability to answer diverse questions, which is rooted in the natural ability to correlate different concepts based on their semantic relationships and decompose difficult problems into sub-tasks. On the contrary, existing visual reasoning methods assume training samples that…

2023

Learning From Unique Perspectives: User-Aware Saliency Modeling

CVPR 2023poster

Everyone is unique. Given the same visual stimuli, people's attention is driven by both salient visual cues and their own inherent preferences. Knowledge of visual preferences not only facilitates understanding of fine-grained attention patterns of diverse users, but also has the potential of benefi…

Cited by 14SourcePDFScholar
2023

Learning Harmonic Molecular Representations on Riemannian Manifold

ICLR 2023poster

Molecular representation learning plays a crucial role in AI-assisted drug discovery research. Encoding 3D molecular structures through Euclidean neural networks has become the prevailing method in the geometric deep learning community. However, the equivariance constraints and message passing in Eu…

2023

No-Regret Learning in Two-Echelon Supply Chain with Unknown Demand Distribution

AISTATS 2023poster

Supply chain management (SCM) has been recognized as an important discipline with applications to many industries, where the two-echelon stochastic inventory model, involving one downstream retailer and one upstream supplier, plays a fundamental role for developing firms’ SCM strategies. In this wor…

Cited by 5SourcePDFScholar
2023

Toward Multi-Granularity Decision-Making: Explicit Visual Reasoning with Hierarchical Knowledge

ICCV 2023poster

Answering visual questions requires the ability to parse visual observations and correlate them with a variety of knowledge. Existing visual question answering (VQA) models either pay little attention to the role of knowledge or do not take into account the granularity of knowledge, e.g., attaching…

Cited by 4PDFcodeScholar
2020

Fantastic Answers and Where to Find Them: Immersive Question-Directed Visual Attention

CVPR 2020poster

While most visual attention studies focus on bottom-up attention with restricted field-of-view, real-life situations are filled with embodied vision tasks. The role of attention is more significant in the latter due to the information overload, and attention to the most important regions is critical…

Cited by 23PDFScholar