← Search

Yuhao Cui

3 accepted papers

2025

Incorporating Dense Knowledge Alignment into Unified Multimodal Representation Models

CVPR 2025poster

Leveraging Large Language Models (LLMs) for text representation has achieved significant success, but the exploration of using Multimodal LLMs (MLLMs) for multimodal representation remains limited. Previous MLLM-based representation studies have primarily focused on unifying the embedding space whil…

Cited by 0SourcePDFScholar
2023

COOP: Decoupling and Coupling of Whole-Body Grasping Pose Generation

ICCV 2023poster

Generating life-like whole-body human grasping has garnered significant attention in the field of computer graphics. Existing works have demonstrated the effectiveness of keyframe-guided motion generation framework, witch focus on modeling the grasping motions of humans in temporal sequence when the…

Cited by 7PDFcodeScholar
2019

Deep Modular Co-Attention Networks for Visual Question Answering

CVPR 2019poster

Visual Question Answering (VQA) requires a fine-grained and simultaneous understanding of both the visual content of images and the textual content of questions. Therefore, designing an effective `co-attention' model to associate key words in questions with key objects in images is central to VQA pe…

Cited by 1110PDFcodeScholar