2024
How to Explore with Belief: State Entropy Maximization in POMDPs
ICML 2024poster
Recent works have studied *state entropy maximization* in reinforcement learning, in which the agent's objective is to learn a policy inducing high entropy over states visitation (Hazan et al., 2019). They typically assume full observability of the state of the system, so that the entropy of the obs…