← Search

Brendan Daniel Tracey

2 accepted papers

2025

Wasserstein Policy Optimization

ICML 2025poster

We introduce Wasserstein Policy Optimization (WPO), an actor-critic algorithm for reinforcement learning in continuous action spaces. WPO can be derived as an approximation to Wasserstein gradient flow over the space of all policies projected into a finite-dimensional parameter space (e.g., the weig…

Cited by 0SourcePDFScholar
2018

On the Information Bottleneck Theory of Deep Learning

ICLR 2018poster

The practical successes of deep neural networks have not been matched by theoretical progress that satisfyingly explains their behavior. In this work, we study the information bottleneck (IB) theory of deep learning, which makes three specific claims: first, that deep networks undergo two distinct p…