← Search

Tom Yan

9 accepted papers

2024

A theoretical case-study of Scalable Oversight in Hierarchical Reinforcement Learning

NeurIPS 2024poster

A key source of complexity in next-generation AI models is the size of model outputs, making it time-consuming to parse and provide reliable feedback on. To ensure such models are aligned, we will need to bolster our understanding of scalable oversight and how to scale up human feedback. To this end…

Cited by 0SourcePDFScholar