2025
Scalable Policy-Based RL Algorithms for POMDPs
NeurIPS 2025poster
The continuous nature of belief states in POMDPs presents significant computational challenges in learning the optimal policy. In this paper, we consider an approach that solves a Partially Observable Reinforcement Learning (PORL) problem by approximating the corresponding POMDP model into a finite-…