← Search

Seongyu Kim

2 accepted papers

2026

Seeing Through Touch: Tactile-Driven Visual Localization of Material Regions

CVPR 2026

We address the problem of tactile localization, where the goal is to identify image regions that share the same material properties as a tactile input. Existing visuo-tactile methods rely on global alignment and thus fail to capture the fine-grained local correspondences required for this task. The

Cited by 0SourcecodeScholar
2025

Seeing Speech and Sound: Distinguishing and Locating Audio Sources in Visual Scenes

CVPR 2025poster

We present a unified model capable of simultaneously grounding both spoken language and non-speech sounds within a visual scene, addressing key limitations in current audio-visual grounding models. Existing approaches are typically limited to handling either speech or non-speech sounds independently…

Cited by 0SourcePDFScholar