← Search

Shan Gao

10 accepted papers

2026

PAMotion: Physics-Aware Motion Generation for Full-Body Interaction with Multiple Objects

CVPR 2026

We present PAMotion, a physics-aware diffusion framework for generating realistic full-body human interactions with multiple objects.Existing diffusion-based methods that jointly synthesize human and object motions often struggle to capture the intricate physical interactions--especially those invol

Cited by 0SourcecodeScholar
2025

Segment Any-Quality Images with Generative Latent Space Enhancement

CVPR 2025poster

Despite their success, Segment Anything Models (SAMs) experience significant performance drops on severely degraded, low-quality images, limiting their effectiveness in real-world scenarios. To address this, we propose GleSAM, which utilizes Generative Latent space Enhancement to boost robustness on…

Cited by 0SourcePDFScholar
2024

APDDv2: Aesthetics of Paintings and Drawings Dataset with Artist Labeled Scores and Comments

NeurIPS 2024poster

Datasets play a pivotal role in training visual models, facilitating the development of abstract understandings of visual features through diverse image samples and multidimensional attributes. However, in the realm of aesthetic evaluation of artistic images, datasets remain relatively scarce. Exist…

2024

P2P: Transforming from Point Supervision to Explicit Visual Prompt for Object Detection and Segmentation

IJCAI 2024poster

Point-supervised vision tasks, including detection and segmentation, aiming to learn a network that transforms from points to pseudo labels, have attracted much attention in recent years. However, the lack of precise object size and boundary annotations in the point-supervised condition results in a…

2024

Paintings and Drawings Aesthetics Assessment with Rich Attributes for Various Artistic Categories

IJCAI 2024poster

Image aesthetic evaluation is a highly prominent research domain in the field of computer vision. In recent years, there has been a proliferation of datasets and corresponding evaluation methodologies for assessing the aesthetic quality of photographic works, leading to the establishment of a relati…

2024

ShapeMatcher: Self-Supervised Joint Shape Canonicalization Segmentation Retrieval and Deformation

CVPR 2024poster

In this paper we present ShapeMatcher a unified self-supervised learning framework for joint shape canonicalization segmentation retrieval and deformation. Given a partially-observed object in an arbitrary pose we first canonicalize the object by extracting point-wise affine invariant features disen…

2023

Tracking without Label: Unsupervised Multiple Object Tracking via Contrastive Similarity Learning

ICCV 2023poster

Unsupervised learning is a challenging task due to the lack of labels. Multiple Object Tracking (MOT), which inevitably suffers from mutual object interference, occlusion, etc., is even more difficult without label supervision. In this paper, we explore the latent consistency of sample features acro…

Cited by 8PDFScholar
2019

Monocular Piecewise Depth Estimation in Dynamic Scenes by Exploiting Superpixel Relations

ICCV 2019poster

In this paper, we propose a novel and specially designed method for piecewise dense monocular depth estimation in dynamic scenes. We utilize spatial relations between neighboring superpixels to solve the inherent relative scale ambiguity (RSA) problem and smooth the depth map. However, directly esti…

Cited by 8PDFScholar