← Search

Shaodi You

18 accepted papers

2025

Automatic Spectral Calibration of Hyperspectral Images: Method, Dataset and Benchmark

CVPR 2025poster

Hyperspectral images (HSI) densely sample the world in both the space and frequency domains and, therefore, are more distinctive than RGB images. Usually, HSI needs to be calibrated to minimize the impact of various illumination conditions. The traditional way to calibrate HSI utilizes a physical re…

2025

Sequential Joint Dependency Aware Human Pose Estimation with State Space Model

AAAI 2025technical

In this paper, we present a sequential joint dependency aware model for monocular 2D-to-3D human pose estimation. While existing estimators leverage the (bi)directional joint dependency with graph convolutions and attention, we further propose to exploit the sequential dependency between joints with…

2024

Atlantis: Enabling Underwater Depth Estimation with Stable Diffusion

CVPR 2024highlight

Monocular depth estimation has experienced significant progress on terrestrial images in recent years thanks to deep learning advancements. But it remains inadequate for underwater scenes primarily due to data scarcity. Given the inherent challenges of light attenuation and backscatter in water acqu…

2024

Learning Content-Enhanced Mask Transformer for Domain Generalized Urban-Scene Segmentation

AAAI 2024technical

Domain-generalized urban-scene semantic segmentation (USSS) aims to learn generalized semantic predictions across diverse urban-scene styles. Unlike generic domain gap challenges, USSS is unique in that the semantic categories are often similar in different urban scenes, while the styles can vary si…

2024

Learning Generalized Segmentation for Foggy-Scenes by Bi-directional Wavelet Guidance

AAAI 2024technical

Learning scene semantics that can be well generalized to foggy conditions is important for safety-crucial applications such as autonomous driving. Existing methods need both annotated clear images and foggy images to train a curriculum domain adaptation model. Unfortunately, these methods can only…

2021

Learning Temporal Consistency for Low Light Video Enhancement From Single Images

CVPR 2021poster

Single image low light enhancement is an important task and it has many practical applications. Most existing methods adopt a single image approach. Although their performance is satisfying on a static single image, we found, however, they suffer serious temporal instability when handling low light…

Cited by 165PDFcodeScholar
2021

Multitask AET With Orthogonal Tangent Regularity for Dark Object Detection

ICCV 2021poster

Dark environment becomes a challenge for computer vision algorithms owing to insufficient photons and undesirable noises. Most of the existing studies tackle this by either targeting human vision for better visual perception or improving the machine vision for specific high-level tasks. In addition,…

Cited by 154PDFcodeScholar
2020

Kinship Identification through Joint Learning using Kinship Verification Ensembles

ECCV 2020poster

Kinship verification is a well-explored task: identifying whether or not two persons are kin. In contrast, kinship identification has been largely ignored so far. Kinship identification aims to further identify the particular type of kinship. An extension to kinship verification run short to properl…

Cited by 20SourcePDFScholar
2020

Unsupervised Learning for Intrinsic Image Decomposition From a Single Image

CVPR 2020poster

Intrinsic image decomposition, which is an essential task in computer vision, aims to infer the reflectance and shading of the scene. It is challenging since it needs to separate one image into two components. To tackle this, conventional methods introduce various priors to constrain the solution, y…

Cited by 133PDFcodeScholar
2019

Classification-Reconstruction Learning for Open-Set Recognition

CVPR 2019poster

Open-set classification is a problem of handling 'unknown' classes that are not contained in the training dataset, whereas traditional classifiers assume that only known classes appear in the test environment. Existing open-set classifiers rely on deep networks trained in a supervised manner on know…

Cited by 531PDFcodeScholar
2018

Single Image Water Hazard Detection using FCN with Reflection Attention Units

ECCV 2018poster

Water bodies, such as puddles and flooded areas, on and off road pose significant risks to autonomous cars. Detecting water from moving camera is a challenging task as water surface is highly refractive, and its appearance varies with viewing angle, surrounding scene, weather conditions. In this pap…

2018

Weakly-Supervised Semantic Segmentation by Iteratively Mining Common Object Features

CVPR 2018poster

Weakly-supervised semantic segmentation under image tags supervision is a challenging task as it directly associates high-level semantic to low-level appearance. To bridge this gap, in this paper, we propose an iterative bottom-up and top-down framework which alternatively expands object regions and…

Cited by 375SourcePDFScholar