← Search

Yujin Sung

2 accepted papers

2026

AxisGuide: Grounding Robot Action Coordinate System in RGB Observations for Robust Visuomotor Manipulation

RSS 2026poster

Visuomotor manipulation policies trained via large-scale behavior cloning have achieved strong semantic scene understanding, yet often fail to reliably execute correct low-level actions under distribution shifts. For example, even in a simple pick-up task with identical scene layouts, camera viewpoi…

Cited by 0SourceScholar
2026

uCLIP: Parameter-Efficient Multilingual Extension of Vision-Language Models with Unpaired Data

AAAI 2026technical

Contrastive Language–Image Pre-training (CLIP) has demonstrated strong generalization across a wide range of visual tasks by leveraging large-scale English–image pairs. However, its extension to low-resource languages remains limited due to the scarcity of high-quality multilingual image–text data.

Cited by 0SourcePDFScholar