← Search

Dongqing Zhang

7 accepted papers

2023

WinCLIP: Zero-/Few-Shot Anomaly Classification and Segmentation

CVPR 2023poster

Visual anomaly classification and segmentation are vital for automating industrial quality inspection. The focus of prior research in the field has been on training custom models for each quality inspection task, which requires task-specific images and annotation. In this paper we move away from thi…

2022

SPot-the-Difference Self-Supervised Pre-training for Anomaly Detection and Segmentation

ECCV 2022poster

"Visual anomaly detection is commonly used in industrial quality inspection. In this paper, we present a new dataset as well as a new self-supervised learning method for ImageNet pre-training to improve anomaly detection and segmentation in 1-class and 2-class 5/10/high-shot training setups. We rele…

2021

Shot Contrastive Self-Supervised Learning for Scene Boundary Detection

CVPR 2021poster

Scenes play a crucial role in breaking the storyline of movies and TV episodes into semantically cohesive parts. However, given their complex temporal structure, finding scene boundaries can be a challenging task requiring large amounts of labeled training data. To address this challenge, we present…

Cited by 87PDFScholar
2018

LQ-Nets: Learned Quantization for Highly Accurate and Compact Deep Neural Networks

ECCV 2018poster

Although weight and activation quantization is an effective approach for Deep Neural Network (DNN) compression and has a lot of potentials to increase inference speed leveraging bit-operations, there is still a noticeable gap in terms of prediction accuracy between the quantized model and the full-p…

2017

ER3: A Unified Framework for Event Retrieval, Recognition and Recounting

CVPR 2017poster

We develop a unified framework for complex event retrieval, recognition and recounting. The framework is based on a compact video representation that exploits the temporal correlations in image features. Our feature alignment procedure identifies and removes the feature redundancies across frames an…

Cited by 28PDFScholar
2017

Neural Aggregation Network for Video Face Recognition

CVPR 2017poster

This paper presents a Neural Aggregation Network (NAN) for video face recognition. The network takes a face video or face image set of a person with a variable number of face images as its input, and produces a compact, fixed-dimension feature representation for recognition. The whole network is com…

Cited by 495PDFScholar
2017

Through the Eustachian Tube and Beyond: A New Miniature Robotic Endoscope to See Into the Middle Ear

RA-L 2017

This paper presents a novel miniature robotic endoscope that is small enough to pass through the Eustachian tube and provide visualization of the middle ear (ME). The device features a miniature bending tip previously conceived of as a small-scale robotic wrist that has been adapted to carry and aim

Cited by 36SourceScholar