← Search

Bryon Tjanaka

5 accepted papers

2026

Discount Model Search for Quality Diversity Optimization in High-Dimensional Measure Spaces

ICLR 2026oral

Quality diversity (QD) optimization searches for a collection of solutions that optimize an objective while attaining diverse outputs of a user-specified, vector-valued measure function. Contemporary QD algorithms are typically limited to low-dimensional measures because high-dimensional measures ar…

Cited by 0SourcecodeScholar
2024

Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning

ICLR 2024spotlight

Training generally capable agents that thoroughly explore their environment and learn new and diverse skills is a long-term goal of robot learning. Quality Diversity Reinforcement Learning (QD-RL) is an emerging research area that blends the best aspects of both fields – Quality Diversity (QD) provi…

Cited by 15SourcePDFScholar
2023

Surrogate Assisted Generation of Human-Robot Interaction Scenarios

CoRL 2023oral

As human-robot interaction (HRI) systems advance, so does the difficulty of evaluating and understanding the strengths and limitations of these systems in different environments and with different users. To this end, previous methods have algorithmically generated diverse scenarios that reveal syste…

Cited by 11SourcecodeScholar
2023

Training Diverse High-Dimensional Controllers by Scaling Covariance Matrix Adaptation MAP-Annealing

RA-L 2023

Pre-training a diverse set of neural network controllers in simulation has enabled robots to adapt online to damage in robot locomotion tasks. However, finding diverse, high-performing controllers requires expensive network training and extensive tuning of a large number of hyperparameters. On the o

Cited by 16SourcecodeScholar
2022

Deep Surrogate Assisted Generation of Environments

NeurIPS 2022accept

Recent progress in reinforcement learning (RL) has started producing generally capable agents that can solve a distribution of complex environments. These agents are typically tested on fixed, human-authored environments. On the other hand, quality diversity (QD) optimization has been proven to be a…