2024
Analyzing Effects of Learning Downstream Tasks on Moral Bias in Large Language Models
COLING 2024main
Pre-training and fine-tuning large language models (LMs) is currently the state-of-the-art methodology for enabling data-scarce downstream tasks. However, the derived models still tend to replicate and perpetuate social biases. To understand this process in more detail, this paper investigates the a…