← Search

Sangha Park

4 accepted papers

2025

RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models

NeurIPS 2025poster

Recent multi-modal large language models (MLLMs) often struggle to generate personalized image captions, even when trained on high-quality captions. In this work, we observe that such limitations persist in existing post-training-based MLLM personalization methods. Specifically, despite being post-t…

Cited by 0SourcecodeScholar
2024

Textual Training for the Hassle-Free Removal of Unwanted Visual Data: Case Studies on OOD and Hateful Image Detection

NeurIPS 2024poster

In our study, we explore methods for detecting unwanted content lurking in visual datasets. We provide a theoretical analysis demonstrating that a model capable of successfully partitioning visual data can be obtained using only textual data. Based on the analysis, we propose Hassle-Free Textual Tra…

2023

On the Powerfulness of Textual Outlier Exposure for Visual OoD Detection

NeurIPS 2023poster

Successful detection of Out-of-Distribution (OoD) data is becoming increasingly important to ensure safe deployment of neural networks. One of the main challenges in OoD detection is that neural networks output overconfident predictions on OoD data, make it difficult to determine OoD-ness of data so…

Cited by 14SourcePDFScholar