← Search

Hyeonsoo Im

2 accepted papers

2026

M^3KG-RAG: Multi-hop Multimodal Knowledge Graph-enhanced Retrieval-Augmented Generation

CVPR 2026

Retrieval-Augmented Generation (RAG) has recently been extended to multimodal settings, connecting multimodal large language models (MLLMs) with vast corpora of external knowledge such as multimodal knowledge graphs (MMKGs). Despite their recent success, multimodal RAG in the audio-visual domain rem

Cited by 0SourceScholar
2025

SelfSplat: Pose-Free and 3D Prior-Free Generalizable 3D Gaussian Splatting

CVPR 2025poster

We propose SelfSplat, a novel 3D Gaussian Splatting model designed to perform pose-free and 3D prior-free generalizable 3D reconstruction from unposed multi-view images. These settings are inherently ill-posed due to the lack of ground-truth data, learned geometric information, and the need to achie…

Cited by 3SourcePDFScholar