Universal Post-Processing Networks for Joint Optimization of Modules in Task-Oriented Dialogue Systems
Post-processing networks (PPNs) are components that modify the outputs of arbitrary modules in task-oriented dialogue systems and are optimized using reinforcement learning (RL) to improve the overall task completion capability of the system. However, previous PPN-based approaches have been limited…