← Search

Xingjia Pan

8 accepted papers

2023

Inter-image Contrastive Consistency for Multi-Person Pose Estimation

AAAI 2023technical

Multi-person pose estimation (MPPE) has achieved impressive progress in recent years. However, due to the large variance of appearances among images or occlusions, the model can hardly learn consistent patterns enough, which leads to severe location jitter and missing issues. In this study, we propo…

Cited by 2SourcePDFScholar
2022

Expanding Low-Density Latent Regions for Open-Set Object Detection

CVPR 2022poster

Modern object detectors have achieved impressive progress under the close-set setup. However, open-set object detection (OSOD) remains challenging since objects of unknown categories are often misclassified to existing known classes. In this work, we propose to identify unknown objects by separating…

Cited by 82PDFcodeScholar
2022

SIOD: Single Instance Annotated per Category per Image for Object Detection

CVPR 2022poster

Object detection under imperfect data receives great attention recently. Weakly supervised object detection (WSOD) suffers from severe localization issues due to the lack of instance-level annotation, while semi-supervised object detection (SSOD) remains challenging led by the inter-image discrepanc…

Cited by 32PDFcodeScholar
2022

SeqTR: A Simple Yet Universal Network for Visual Grounding

ECCV 2022poster

"In this paper, we propose a simple yet universal network termed SeqTR for visual grounding tasks, e.g., phrase localization, referring expression comprehension (REC) and segmentation (RES). The canonical paradigms for visual grounding often require substantial expertise in designing network archite…

2022

StyTr2: Image Style Transfer With Transformers

CVPR 2022poster

The goal of image style transfer is to render an image with artistic features guided by a style reference while maintaining the original content. Owing to the locality in convolutional neural networks (CNNs), extracting and maintaining the global information of input images is difficult. Therefore,…

Cited by 379PDFcodeScholar
2021

TS-CAM: Token Semantic Coupled Attention Map for Weakly Supervised Object Localization

ICCV 2021poster

Weakly supervised object localization (WSOL) is a challenging problem when given image category labels but requires to learn object localization models. Optimizing a convolutional neural network (CNN) for classification tends to activate local discriminative regions while ignoring complete object ex…

Cited by 254PDFcodeScholar
2021

Unveiling the Potential of Structure Preserving for Weakly Supervised Object Localization

CVPR 2021poster

Weakly supervised object localization (WSOL) remains an open problem due to the deficiency of finding object extent information using a classification network. While prior works struggle to localize objects by various spatial regularization strategies, we argue that how to extract object structural…

Cited by 110PDFcodeScholar
2020

Dynamic Refinement Network for Oriented and Densely Packed Object Detection

CVPR 2020oral

Object detection has achieved remarkable progress in the past decade. However, the detection of oriented and densely packed objects remains challenging because of following inherent reasons: (1) receptive fields of neurons are all axis-aligned and of the same shape, whereas objects are usually of di…

Cited by 411PDFcodeScholar