← Search

Jiaxin Lu

8 accepted papers

2025

HUMOTO: A 4D Dataset of Mocap Human Object Interactions

ICCV 2025poster

We present Human Motions with Objects (HUMOTO), a high-fidelity dataset of human-object interactions for motion generation, computer vision, and robotics applications. Featuring 735 sequences (7,875 seconds at 30 fps), HUMOTO captures interactions with 63 precisely modeled objects and 72 articulated…

2025

Learning Structured Universe Graph with Outlier OOD Detection for Partial Matching

ICLR 2025poster

Partial matching is a kind of graph matching where only part of two graphs can be aligned. This problem is particularly important in computer vision applications, where challenges like point occlusion or annotation errors often occur when labeling key points. Previous work has often conflated point…

Cited by 0SourcePDFScholar
2025

VQAGuider: Guiding Multimodal Large Language Models to Answer Complex Video Questions

ACL 2025long

Complex video question-answering (VQA) requires in-depth understanding of video contents including object and action recognition as well as video classification and summarization, which exhibits great potential in emerging applications in education and entertainment, etc. Multimodal large language m…

2024

M3C: A Framework towards Convergent, Flexible, and Unsupervised Learning of Mixture Graph Matching and Clustering

ICLR 2024poster

Existing graph matching methods typically assume that there are similar structures between graphs and they are matchable. This work addresses a more realistic scenario where graphs exhibit diverse modes, requiring graph grouping before or along with matching, a task termed mixture graph matching and…

Cited by 1SourcePDFScholar