Markov Balance Satisfaction Improves Performance in Strictly Batch Offline Imitation Learning
Imitation learning (IL) is notably effective for robotic tasks where directly programming behaviors or defining optimal control costs is challenging. In this work, we address a scenario where the imitator relies solely on observed behavior and cannot make environmental interactions during learning.…