← Search

Myunsoo Kim

5 accepted papers

2026

Beyond RAG vs. Long-Context: Learning Distraction-Aware Retrieval for Efficient Knowledge Grounding

ICLR 2026poster

Retrieval-Augmented Generation (RAG) is a framework for grounding Large Language Models (LLMs) in external, up-to-date information. However, recent advancements in context window size allow LLMs to process inputs of up to 128K tokens or more, offering an alternative strategy: supplying the full docu…

Cited by 0SourceScholar
2026

FALCON: False-Negative Aware Learning of Contrastive Negatives in Vision-Language Alignment

CVPR 2026

False negatives pose a critical challenge in vision-language pretraining (VLP) due to the many-to-many correspondence between images and texts in large-scale datasets. These false negatives introduce conflicting supervision signals that degrade the learned embedding space and diminish the effectiven

Cited by 0SourcecodeScholar
2025

Adaptive Non-Uniform Timestep Sampling for Accelerating Diffusion Model Training

CVPR 2025poster

As a highly expressive generative model, diffusion models have demonstrated exceptional success across various domains, including image generation, natural language processing, and combinatorial optimization. However, as data distributions grow more complex, training these models to convergence beco…

Cited by 0SourcePDFScholar
2025

NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations

ICML 2025poster

Intelligent agents are able to make decisions based on different levels of granularity and duration. Recent advances in skill learning enabled the agent to solve complex, long-horizon tasks by effectively guiding the agent in choosing appropriate skills. However, the practice of using fixed-length s…

2025

Rethinking DPO: The Role of Rejected Responses in Preference Misalignment

EMNLP 2025

Direct Preference Optimization (DPO) is a simple and efficient framework that has attracted substantial attention. However, it often struggles to meet its primary objectives—increasing the generation probability of chosen responses while reducing that of rejected responses—due to the dominant influe