← Search

Qifeng Zhou

2 accepted papers

2026

Hyperbolic Gramian Volumes for Multimodal Alignment

CVPR 2026

Multimodal contrastive learning typically relies on pairwise similarities for alignment, but recent work has shown that Gramian volumes can capture higher-order correlations across modalities. However, Euclidean Gramian volumes suffer from volume collapse under L2 normalization, concentrating near u

Cited by 0SourceScholar
2024

A Hierarchical Network for Multimodal Document-Level Relation Extraction

AAAI 2024technical

Document-level relation extraction aims to extract entity relations that span across multiple sentences. This task faces two critical issues: long dependency and mention selection. Prior works address the above problems from the textual perspective, however, it is hard to handle these problems solel…