← Search

Qiaoyong Zhong

5 accepted papers

2025

VQCounter: Designing Visual Prompt Queue for Accurate Open-World Counting

IJCAI 2025

Class-agnostic counting enables enumerating arbitrary object classes beyond those seen during training. Recent studies attempted to exploit the potential of visual foundation models such as GroundingDINO. Despite the considerable progress, we observe certain shortcomings, including the limited diver

Cited by 0SourcePDFScholar
2023

Towards Deployment-Efficient and Collision-Free Multi-Agent Path Finding (Student Abstract)

AAAI 2023technical

Multi-agent pathfinding (MAPF) is essential to large-scale robotic coordination tasks. Planning-based algorithms show their advantages in collision avoidance while avoiding exponential growth in the number of agents. Reinforcement-learning (RL)-based algorithms can be deployed efficiently but cannot…

Cited by 0SourcePDFScholar
2022

Topology-Aware Convolutional Neural Network for Efficient Skeleton-Based Action Recognition

AAAI 2022technical

In the context of skeleton-based action recognition, graph convolutional networks (GCNs) have been rapidly developed, whereas convolutional neural networks (CNNs) have received less attention. One reason is that CNNs are considered poor in modeling the irregular skeleton topology. To alleviate this…

2021

Divide-and-Assemble: Learning Block-Wise Memory for Unsupervised Anomaly Detection

ICCV 2021poster

Reconstruction-based methods play an important role in unsupervised anomaly detection in images. Ideally, we expect a perfect reconstruction for normal samples and poor reconstruction for abnormal samples. Since the generalizability of deep neural networks is difficult to control, existing models su…

Cited by 193PDFScholar
2019

Collaborative Spatiotemporal Feature Learning for Video Action Recognition

CVPR 2019poster

Spatiotemporal feature learning is of central importance for action recognition in videos. Existing deep neural network models either learn spatial and temporal features independently (C2D) or jointly with unconstrained parameters (C3D). In this paper, we propose a novel neural operation which encod…

Cited by 131PDFcodeScholar