Concurrent Reinforcement Learning with Aggregated States via Randomized Least Squares Value Iteration
Designing learning agents that explore efficiently in a complex environment has been widely recognized as a fundamental challenge in reinforcement learning. While a number of works have demonstrated the effectiveness of techniques based on randomized value functions on a single agent, it remains un…