2025
Think Globally, Group Locally: Evaluating LLMs Using Multi-Lingual Word Grouping Games
EMNLP 2025
Large language models (LLMs) can exhibit biases in reasoning capabilities due to linguistic modality, performing better on tasks in one language versus another, even with similar content. Most previous works evaluate this through reasoning tasks where reliance on strategies or knowledge can ensure s