← Search

Hongjun Zhou

6 accepted papers

2025

Self-Organised Sequential Multi-Agent Reinforcement Learning for Closely Cooperation Tasks

RA-L 2025

Cooperative tasks are common in multi-agent systems, with closely cooperative tasks being a special case of this, where a change in the state of the environment requires multiple agents to perform a specific operation at the same time. Take a box-pushing task as an example, the box is heavy and requ

Cited by 0SourceScholar
2024

Closely Cooperative Multi-Agent Reinforcement Learning Based on Intention Sharing and Credit Assignment

RA-L 2024

Collaborative tasks are important in multi-agent systems. Multi-agent reinforcement learning is a commonly used technique for solving multi-agent cooperative policy learning. The closely collaborative task is a special but common case within cooperative tasks, where the change in the environmental s

Cited by 2SourceScholar
2023

GAN-Based Editable Movement Primitive From High-Variance Demonstrations

RA-L 2023

Movement Primitive (MP) is a promising Learning from Demonstration (LfD) framework, which is commonly used to learn movements from human demonstrations and adapt the learned movements to new task scenes. A major goal of MP research is to improve the adaptability of MP to various target positions and

Cited by 4SourceScholar
2023

Goal-Conditioned Reinforcement Learning With Disentanglement-Based Reachability Planning

RA-L 2023

Goal-Conditioned Reinforcement Learning (GCRL) can enable agents to spontaneously set diverse goals to learn a set of skills. Despite the excellent works proposed in various fields, reaching distant goals in temporally extended tasks remains a challenge for GCRL. Current works tackled this problem b

Cited by 6SourceScholar
2022

Weakly Supervised Disentangled Representation for Goal-Conditioned Reinforcement Learning

RA-L 2022

Goal-conditioned reinforcement learning is a crucial yet challenging algorithm which enables agents to achieve multiple user-specified goals when learning a set of skills in a dynamic environment. However, it typically requires millions of the environmental interactions explored by agents, which is

Cited by 7SourceScholar