2025
An Efficient Dialogue Policy Agent with Model-Based Causal Reinforcement Learning
COLING 2025main
Dialogue policy trains an agent to select dialogue actions frequently implemented via deep reinforcement learning (DRL). The model-based reinforcement methods built a world model to generate simulated data to alleviate the sample inefficiency. However, traditional world model methods merely consider…