← Search

Yayun Qi

2 accepted papers

2026

What to Trust? A Trust-aware Knowledge-guided Method for Zero-shot Object State Understanding in Videos

AAAI 2026technical

Object state understanding aims at recognizing the co-occurrence and transitions of multiple object states in videos. While learning from videos handles seen object states well, it struggles with novel ones. We address this task in a zero-shot setting by extracting state-specific knowledge from pre-

Cited by 0SourcePDFScholar
2024

Relational Distant Supervision for Image Captioning without Image-Text Pairs

AAAI 2024technical

Unsupervised image captioning aims to generate descriptions of images without relying on any image-sentence pairs for training. Most existing works use detected visual objects or concepts as bridge to connect images and texts. Considering that the relationship between objects carries more informatio…

Cited by 5SourcePDFScholar