← Search

Andrea Vallone

2 accepted papers

2024

Rule Based Rewards for Language Model Safety

NeurIPS 2024poster

Reinforcement learning based fine-tuning of large language models (LLMs) on human preferences has been shown to enhance both their capabilities and safety behavior. However, in cases related to safety, without precise instructions to human annotators, the data collected may cause the model to beco…

2022

Danish Airs and Grounds: A Dataset for Aerial-to-Street-Level Place Recognition and Localization

RA-L 2022

Place recognition and visual localization are particularly challenging in wide baseline configurations. In this letter, we contribute with the <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">Danish Airs and Grounds</i> (DAG) dataset, a large collecti

Cited by 12SourceScholar