← Search

Gregory Serapio-García

2 accepted papers

2024

GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives

NAACL 2024long

Human annotation plays a core role in machine learning — annotations for supervised models, safety guardrails for generative models, and human feedback for reinforcement learning, to cite a few avenues. However, the fact that many of these human annotations are inherently subjective is often overloo…

2024

Moral Foundations of Large Language Models

EMNLP 2024main

Moral foundations theory (MFT) is a psychological assessment tool that decomposes human moral reasoning into five factors, including care/harm, liberty/oppression, and sanctity/degradation (Graham et al., 2009). People vary in the weight they place on these dimensions when making moral decisions, in…