2025
PoSum-Bench: Benchmarking Position Bias in LLM-based Conversational Summarization
EMNLP 2025
Large language models (LLMs) are increasingly used for zero-shot conversation summarization, but often exhibit positional bias—tending to overemphasize content from the beginning or end of a conversation while neglecting the middle. To address this issue, we introduce PoSum-Bench, a comprehensive be