2018
Benchmarking Uncertainty Estimates with Deep Reinforcement Learning for Dialogue Policy Optimisation
ICASSP 2018accepted
In statistical dialogue management, the dialogue manager learns a policy that maps a belief state to an action for the system to perform. Efficient exploration is key to successful policy optimisation. Current deep reinforcement learning methods are very promising but rely on ε-greedy exploration, t…