2021
Width-based Lookaheads with Learnt Base Policies and Heuristics Over the Atari-2600 Benchmark
NeurIPS 2021spotlight
We propose new width-based planning and learning algorithms inspired from a careful analysis of the design decisions made by previous width-based planners. The algorithms are applied over the Atari-2600 games and our best performing algorithm, Novelty guided Critical Path Learning (N-CPL), outperfor…