2023
Taming Continuous Posteriors for Latent Variational Dialogue Policies
AAAI 2023technical
Utilizing amortized variational inference for latent-action reinforcement learning (RL) has been shown to be an effective approach in Task-oriented Dialogue (ToD) systems for optimizing dialogue success.Until now, categorical posteriors have been argued to be one of the main drivers of performance.…