← Search

Dennis Lee

5 accepted papers

2021

Mathematical Reasoning via Self-supervised Skip-tree Training

ICLR 2021spotlight

We demonstrate that self-supervised language modeling applied to mathematical formulas enables logical reasoning. To measure the logical reasoning abilities of language models, we formulate several evaluation (downstream) tasks, such as inferring types, suggesting missing assumptions and completing…

Cited by 59SourcePDFScholar
2020

Model-based Reinforcement Learning for Decentralized Multiagent Rendezvous

CoRL 2020

Collaboration requires agents to align their goals on the fly. Underlying the human ability to align goals with other agents is their ability to predict the intentions of others and actively update their own plans. We propose hierarchical predictive planning (HPP), a model-based reinforcement learni

Cited by 0SourcePDFScholar
2019

ProMP: Proximal Meta-Policy Search

ICLR 2019poster

Credit assignment in Meta-reinforcement learning (Meta-RL) is still poorly understood. Existing methods either neglect credit assignment to pre-adaptation behavior or implement it naively. This leads to poor sample-efficiency during meta-training as well as ineffective task identification strategies…

2018

Deep Imitation Learning for Complex Manipulation Tasks from Virtual Reality Teleoperation

ICRA 2018poster

Imitation learning is a powerful paradigm for robot skill acquisition. However, obtaining demonstrations suitable for learning a policy that maps from raw pixels to actions can be challenging. In this paper we describe how consumer-grade Virtual Reality headsets and hand tracking hardware can be use…

Cited by 895SourceScholar