← Search

Guiqian Zhu

2 accepted papers

2025

Navigating the Unseen: Zero-shot Scene Graph Generation via Capsule-Based Equivariant Features

CVPR 2025poster

In scene graph generation (SGG), the accurate prediction of unseen triples is essential for its effectiveness in downstream vision-language tasks. We hypothesize that the predicates of unseen triples can be viewed as transformations of seen predicates in feature space, and the essence of the zero-sh…

Cited by 0SourcePDFScholar
2024

TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling

COLING 2024main

As a cross-modal task, visual storytelling aims to generate a story for an ordered image sequence automatically. Different from the image captioning task, visual storytelling requires not only modeling the relationships between objects in the image but also mining the connections between adjacent im…

Cited by 1SourcePDFScholar