← Search

Nafiseh Nikeghbal

2 accepted papers

2025

CoBia: Constructed Conversations Can Trigger Otherwise Concealed Societal Biases in LLMs

EMNLP 2025

Improvements in model construction, including fortified safety guardrails, allow Large language models (LLMs) to increasingly pass standard safety checks. However, LLMs sometimes slip into revealing harmful behavior, such as expressing racist viewpoints, during conversations. To analyze this systema

2025

MEXA: Multilingual Evaluation of English-Centric LLMs via Cross-Lingual Alignment

ACL 2025finding

English-centric large language models (LLMs) often show strong multilingual capabilities. However, their multilingual performance remains unclear and is under-evaluated for many other languages. Most benchmarks for multilinguality focus on classic NLP tasks or cover a minimal number of languages. We…