← Search

Gefei Gu

2 accepted papers

2025

Ref-Long: Benchmarking the Long-context Referencing Capability of Long-context Language Models

ACL 2025long

Long-context language models (LCLMs) have exhibited impressive capabilities in long-context understanding tasks. Among these, long-context referencing—a crucial task that requires LCLMs to attribute items of interest to specific parts of long-context data—remains underexplored. To bridge this gap, t…

2024

TAIL: A Toolkit for Automatic and Realistic Long-Context Large Language Model Evaluation

EMNLP 2024system demonstrations

As long-context large language models (LLMs) are attracting increasing attention for their ability to handle context windows exceeding 128k tokens, the need for effective evaluation methods for these models becomes critical.Existing evaluation methods, however, fall short: needle-in-a-haystack (NIAH…