← Search

Jiangxing Wang

6 accepted papers

2024

Language Model Adaption for Reinforcement Learning with Natural Language Action Space

ACL 2024long

Reinforcement learning with natural language action space often suffers from the curse of dimensionality due to the combinatorial nature of the natural language. Previous research leverages pretrained language models to capture action semantics and reduce the size of the action space. However, since…

2023

More Centralized Training, Still Decentralized Execution: Multi-Agent Conditional Policy Factorization

ICLR 2023poster

In cooperative multi-agent reinforcement learning (MARL), combining value decomposition with actor-critic enables agents to learn stochastic policies, which are more suitable for the partially observable environment. Given the goal of learning local policies that enable decentralized execution, agen…