← Search

Andrés Cotton

1 accepted papers

2024

The Greatest Good Benchmark: Measuring LLMs’ Alignment with Utilitarian Moral Dilemmas

EMNLP 2024main

The question of how to make decisions that maximise the well-being of all persons is very relevant to design language models that are beneficial to humanity and free from harm. We introduce the Greatest Good Benchmark to evaluate the moral judgments of LLMs using utilitarian dilemmas. Our analysis a…