2024
LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
ACL 2024long
Although large language models (LLMs) demonstrate impressive performance for many language tasks, most of them can only handle texts a few thousand tokens long, limiting their applications on longer sequence inputs, such as books, reports, and codebases. Recent works have proposed methods to improve…