← Search

Sahil Badyal

3 accepted papers

2020

Multiagent Rollout and Policy Iteration for POMDP with Application to Multi-Robot Repair Problems

CoRL 2020

In this paper we consider infinite horizon discounted dynamic programming problems with finite state and control spaces, partial state observations, and a multiagent structure. We discuss and compare algorithms that simultaneously or sequentially optimize the agents’ controls by using multistep look

Cited by 0SourcePDFScholar
2020

Reinforcement Learning for POMDP: Partitioned Rollout and Policy Iteration With Application to Autonomous Sequential Repair Problems

RA-L 2020

In this letter we consider infinite horizon discounted dynamic programming problems with finite state and control spaces, and partial state observations. We discuss an algorithm that uses multistep lookahead, truncated rollout with a known base policy, and a terminal cost function approximation. Thi

Cited by 37SourceScholar