← Search

Oliver Hayman

1 accepted papers

2024

Goodhart's Law in Reinforcement Learning

ICLR 2024poster

Implementing a reward function that perfectly captures a complex task in the real world is impractical. As a result, it is often appropriate to think of the reward function as a *proxy* for the true objective rather than as its definition. We study this phenomenon through the lens of *Goodhart’s law…

Cited by 13SourcePDFScholar