← Search

Rahma Chaabouni

6 accepted papers

2024

Countering Reward Over-Optimization in LLM with Demonstration-Guided Reinforcement Learning

ACL 2024findings

While reinforcement learning (RL) has been proven essential for tuning large language models (LLMs), it can lead to reward over-optimization (ROO). Existing approaches address ROO by adding KL regularization, requiring computationally expensive hyperparameter tuning. Additionally, KL regularization…

2024

Memory Consolidation Enables Long-Context Video Understanding

ICML 2024spotlight

Most transformer-based video encoders are limited to short temporal contexts due to their quadratic complexity. While various attempts have been made to extend this context, this has often come at the cost of both conceptual and computational complexity. We propose to instead re-purpose existing pre…

Cited by 25SourcePDFScholar
2022

Emergent Communication at Scale

ICLR 2022spotlight

Emergent communication aims for a better understanding of human language evolution and building more efficient representations. We posit that reaching these goals will require scaling up, in contrast to a significant amount of literature that focuses on setting up small-scale problems to tease out d…

2020

Entropy Minimization In Emergent Languages

ICML 2020poster

There is growing interest in studying the languages that emerge when neural agents are jointly trained to solve tasks requiring communication through a discrete channel. We investigate here the information-theoretic complexity of such languages, focusing on the basic two-agent, one-exchange setup. W…

2019

Anti-efficient encoding in emergent communication

NeurIPS 2019poster

Despite renewed interest in emergent language simulations with neural networks, little is known about the basic properties of the induced code, and how they compare to human language. One fundamental characteristic of the latter, known as Zipf's Law of Abbreviation (ZLA), is that more freque…