2026
Do It for HER: First-Order Temporal Logic Reward Specification in Reinforcement Learning
AAAI 2026technical
In this work, we propose a novel framework for the logical specification of non-Markovian rewards in Markov Decision Processes (MDPs) with large state spaces. Our approach leverages Linear Temporal Logic Modulo Theories over finite traces (LTLfMT), a more expressive extension of classical temporal l