2025
MDP Geometry, Normalization and Reward Balancing Solvers
AISTATS 2025poster
We present a new geometric interpretation of Markov Decision Processes (MDPs) with a natural normalization procedure that allows us to adjust the value function at each state without altering the advantage of any action with respect to any policy. This advantage-preserving transformation of the MDP…