← Search

Junghyuk Yeom

2 accepted papers

2024

Exclusively Penalized Q-learning for Offline Reinforcement Learning

NeurIPS 2024spotlight

Constraint-based offline reinforcement learning (RL) involves policy constraints or imposing penalties on the value function to mitigate overestimation errors caused by distributional shift. This paper focuses on a limitation in existing offline RL methods with penalized value function, indicating t…

Cited by 2SourcePDFScholar
2024

FoX: Formation-Aware Exploration in Multi-Agent Reinforcement Learning

AAAI 2024technical

Recently, deep multi-agent reinforcement learning (MARL) has gained significant popularity due to its success in various cooperative multi-agent tasks. However, exploration still remains a challenging problem in MARL due to the partial observability of the agents and the exploration space that can g…