2022
Efficient Dialog Policy Learning by Reasoning with Contextual Knowledge
AAAI 2022technical
Goal-oriented dialog policy learning algorithms aim to learn a dialog policy for selecting language actions based on the current dialog state. Deep reinforcement learning methods have been used for dialog policy learning. This work is motivated by the observation that, although dialog is a domain wi…