2024
Interpretable Concept Bottlenecks to Align Reinforcement Learning Agents
NeurIPS 2024poster
Goal misalignment, reward sparsity and difficult credit assignment are only a few of the many issues that make it difficult for deep reinforcement learning (RL) agents to learn optimal policies. Unfortunately, the black-box nature of deep neural networks impedes the inclusion of domain experts for…