2022
A Versatile Adaptive Curriculum Learning Framework for Task-oriented Dialogue Policy Learning
NAACL 2022findings
Training a deep reinforcement learning-based dialogue policy with brute-force random sampling is costly. A new training paradigm was proposed to improve learning performance and efficiency by combining curriculum learning. However, attempts in the field of dialogue policy are very limited due to the…