← Search

Xiaofeng Ji

2 accepted papers

2024

Multi-Modal Prompting for Open-Vocabulary Video Visual Relationship Detection

AAAI 2024technical

Open-vocabulary video visual relationship detection aims to extend video visual relationship detection beyond annotated categories by detecting unseen relationships between objects in videos. Recent progresses in open-vocabulary perception, primarily driven by large-scale image-text pre-trained mod…

2022

Adaptive Image-to-Video Scene Graph Generation via Knowledge Reasoning and Adversarial Learning

AAAI 2022technical

Scene graph in a video conveys a wealth of information about objects and their relationships in the scene, thus benefiting many downstream tasks such as video captioning and visual question answering. Existing methods of scene graph generation require large-scale training videos annotated with objec…

Cited by 1SourcePDFScholar