← Search

Ruozhen He

3 accepted papers

2024

Improved Visual Grounding through Self-Consistent Explanations

CVPR 2024poster

Vision-and-language models trained to match images with text can be combined with visual explanation methods to point to the locations of specific objects in an image. Our work shows that the localization --"grounding'"-- abilities of these models can be further improved by finetuning for self-consi…

Cited by 15SourcePDFScholar
2023

Efficient Mirror Detection via Multi-Level Heterogeneous Learning

AAAI 2023technical

We present HetNet (Multi-level Heterogeneous Network), a highly efficient mirror detection network. Current mirror detection methods focus more on performance than efficiency, limiting the real-time applications (such as drones). Their lack of efficiency is aroused by the common design of adopting h…

2023

Weakly-Supervised Camouflaged Object Detection with Scribble Annotations

AAAI 2023technical

Existing camouflaged object detection (COD) methods rely heavily on large-scale datasets with pixel-wise annotations. However, due to the ambiguous boundary, annotating camouflage objects pixel-wisely is very time-consuming and labor-intensive, taking ~60mins to label one image. In this paper, we pr…