← Search

Ravi Srinivasan

2 accepted papers

2022

Cross-Domain Detection of GPT-2-Generated Technical Text

NAACL 2022long

Machine-generated text presents a potential threat not only to the public sphere, but also to the scientific enterprise, whereby genuine research is undermined by convincing, synthetic text. In this paper we examine the problem of detecting GPT-2-generated technical research text. We first consider…

2020

Safe Imitation Learning via Fast Bayesian Reward Inference from Preferences

ICML 2020poster

Bayesian reward learning from demonstrations enables rigorous safety and uncertainty analysis when performing imitation learning. However, Bayesian reward learning methods are typically computationally intractable for complex control problems. We propose Bayesian Reward Extrapolation (Bayesian REX),…