2025
NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens
ICLR 2025poster
Recent advancements in Large Language Models (LLMs) have pushed the boundaries of natural language processing, especially in long-context understanding. However, the evaluation of these models' long-context abilities remains a challenge due to the limitations of current benchmarks. To address this g…