Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
Evaluating learned robot control policies to determine their performance costs the experimenter time and effort. As robots become more capable in accomplishing diverse tasks, evaluating across all these tasks becomes more difficult as it is impractical to test every policy on every task multiple tim…