2024
Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA
EMNLP 2024main
Long-context modeling capabilities of Large Language Models (LLMs) have garnered widespread attention, leading to the emergence of LLMs with ultra-context windows. Meanwhile, benchmarks for evaluating long-context language models are gradually catching up. However, existing benchmarks employ irrelev…