2025
Exploring and Mitigating Implicit Bias in Large Language Models: A Cross-Domain Evaluation Framework
AAAI 2025technical
This paper investigates implicit biases in large language models (LLMs) triggered by subtle contextual cues. Through experiments, the study examines how these biases influence model outputs in domains such as healthcare and hiring. A framework for mitigating stereotype reinforcement is proposed, alo…