← Search

Hengcan Shi

8 accepted papers

2025

DrVideo: Document Retrieval Based Long Video Understanding

CVPR 2025poster

Most of the existing methods for video understanding primarily focus on videos only lasting tens of seconds, with limited exploration of techniques for handling long videos. The increased number of frames in long videos poses two main challenges: difficulty in locating key information and performing…

2024

JRDB-PanoTrack: An Open-world Panoptic Segmentation and Tracking Robotic Dataset in Crowded Human Environments

CVPR 2024poster

Autonomous robot systems have attracted increasing research attention in recent years where environment understanding is a crucial step for robot navigation human-robot interaction and decision. Real-world robot systems usually collect visual data from multiple sensors and are required to recognize…

Cited by 2SourcePDFScholar
2023

Accurate and Real-Time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network

RA-L 2023

Efficiently and accurately detecting people from 3D point cloud data is of great importance in many robotic and autonomous driving applications. This fundamental perception task is still very challenging due to <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/199

Cited by 29SourcecodeScholar
2022

ProposalCLIP: Unsupervised Open-Category Object Proposal Generation via Exploiting CLIP Cues

CVPR 2022poster

Object proposal generation is an important and fundamental task in computer vision. In this paper, we propose ProposalCLIP, a method towards unsupervised open-category object proposal generation. Unlike previous works which require a large number of bounding box annotations and/or can only generate…

Cited by 70PDFScholar
2019

Scene Parsing via Integrated Classification Model and Variance-Based Regularization

CVPR 2019poster

Scene Parsing is a challenging task in computer vision, which can be formulated as a pixel-wise classification problem. Existing deep-learning-based methods usually use one general classifier to recognize all object categories. However, the general classifier easily makes some mistakes in dealing wi…

Cited by 14PDFcodeScholar
2018

Key-Word-Aware Network for Referring Expression Image Segmentation

ECCV 2018poster

Referring expression image segmentation aims to segment out the object referred by a natural language query expression. Without considering the specific properties of visual and textual information, existing works usually deal with this task by directly feeding a foreground/background classifier wit…

Cited by 214SourcePDFScholar