2025
BraVE: Offline Reinforcement Learning for Discrete Combinatorial Action Spaces
NeurIPS 2025poster
Offline reinforcement learning in high-dimensional, discrete action spaces is challenging due to the exponential scaling of the joint action space with the number of sub-actions and the complexity of modeling sub-action dependencies. Existing methods either exhaustively evaluate the action space, ma…