← Search

Ekin Akyurek

4 accepted papers

2023

RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs

ACL 2023long

Despite their unprecedented success, even the largest language models make mistakes. Similar to how humans learn and improve using feedback, previous work proposed providing language models with natural language feedback to guide them in repairing their outputs. Because human-generated critiques are…

2022

Towards Tracing Knowledge in Language Models Back to the Training Data

EMNLP 2022finding

Language models (LMs) have been shown to memorize a great deal of factual knowledge contained in their training data. But when an LM generates an assertion, it is often difficult to determine where it learned this information and whether it is true. In this paper, we propose the problem of fact trac…