← Search

Bumsoo Park

3 accepted papers

2025

Learning to Better Search with Language Models via Guided Reinforced Self-Training

NeurIPS 2025poster

While language models have shown remarkable performance across diverse tasks, they still encounter challenges in complex reasoning scenarios. Recent research suggests that language models trained on linearized search traces toward solutions, rather than solely on the final solutions, exhibit improve…

Cited by 0SourcecodeScholar
2023

Discovering Hierarchical Achievements in Reinforcement Learning via Contrastive Learning

NeurIPS 2023poster

Discovering achievements with a hierarchical structure in procedurally generated environments presents a significant challenge. This requires an agent to possess a broad range of abilities, including generalization and long-term reasoning. Many prior methods have been built upon model-based or hiera…