← Search

JunHyeok Oh

4 accepted papers

2026

TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning

ICML 2026poster

The design of environments plays a critical role in shaping the development and evaluation of cooperative multi-agent reinforcement learning (MARL) algorithms. While existing benchmarks highlight critical challenges, they often lack the modularity required to design custom evaluation scenarios. We i…

Cited by 0SourceScholar
2025

Iterative Prompt Refinement for Safer Text-to-Image Generation

EMNLP 2025

Text-to-Image (T2I) models have made remarkable progress in generating images from text prompts, but their output quality and safety still depend heavily on how prompts are phrased. Existing safety methods typically refine prompts using large language models (LLMs), but they overlook the images prod

2025

Prior-Guided Diffusion Planning for Offline Reinforcement Learning

NeurIPS 2025poster

Diffusion models have recently gained prominence in offline reinforcement learning due to their ability to effectively learn high-performing, generalizable policies from static datasets. Diffusion-based planners facilitate long-horizon decision-making by generating high-quality trajectories through…

Cited by 13SourcecodeScholar
2025

Rethinking DPO: The Role of Rejected Responses in Preference Misalignment

EMNLP 2025

Direct Preference Optimization (DPO) is a simple and efficient framework that has attracted substantial attention. However, it often struggles to meet its primary objectives—increasing the generation probability of chosen responses while reducing that of rejected responses—due to the dominant influe