← Search

Gyuri Szarvas

1 accepted papers

2023

Taming Continuous Posteriors for Latent Variational Dialogue Policies

AAAI 2023technical

Utilizing amortized variational inference for latent-action reinforcement learning (RL) has been shown to be an effective approach in Task-oriented Dialogue (ToD) systems for optimizing dialogue success.Until now, categorical posteriors have been argued to be one of the main drivers of performance.…

Cited by 2SourcePDFScholar