← Search

Sebastian Sztwiertnia

1 accepted papers

2024

Interpretable Concept Bottlenecks to Align Reinforcement Learning Agents

NeurIPS 2024poster

Goal misalignment, reward sparsity and difficult credit assignment are only a few of the many issues that make it difficult for deep reinforcement learning (RL) agents to learn optimal policies. Unfortunately, the black-box nature of deep neural networks impedes the inclusion of domain experts for…