← Search

Byeongguk Jeon

4 accepted papers

2026

Q-Flow: Stable and Expressive Reinforcement Learning with Flow-based Policy

ICML 2026poster

There is growing interest in utilizing flow-based models as decision-making policies in reinforcement learning due to their high expressive capacity. However, effectively leveraging this expressivity for value maximization remains challenging, as naive gradient-based optimization requires backpropag…

Cited by 0SourceScholar
2025

Ask Optimal Questions: Aligning Large Language Models with Retriever’s Preference in Conversation

NAACL 2025findings

Conversational search, unlike single-turn retrieval tasks, requires understanding the current question within a dialogue context. The common approach of rewrite-then-retrieve aims to decontextualize questions to be self-sufficient for off-the-shelf retrievers, but most existing methods produce sub-o…

2025

Latent Action Pretraining from Videos

ICLR 2025poster

We introduce Latent Action Pretraining for general Action models (LAPA), the first unsupervised method for pretraining Vision-Language-Action (VLA) models without ground-truth robot action labels. Existing Vision-Language-Action models require action labels typically collected by human teleoperators…

Cited by 20SourcePDFScholar
2023

Tree of Clarifications: Answering Ambiguous Questions with Retrieval-Augmented Large Language Models

EMNLP 2023short main

Questions in open-domain question answering are often ambiguous, allowing multiple interpretations. One approach to handling them is to identify all possible interpretations of the ambiguous question (AQ) and to generate a long-form answer addressing them all, as suggested by Stelmakh et al., (2022…

Cited by 0SourcecodeScholar