2015
Correct-by-synthesis reinforcement learning with temporal logic constraints
IROS 2015poster
We consider a problem on the synthesis of optimal reactive controllers with an a priori unknown performance criterion while satisfying a given temporal logic specification through the interaction with an uncontrolled environment. We decouple the problem into two sub-problems. First, we extract a (ma…