2024
∞Bench: Extending Long Context Evaluation Beyond 100K Tokens
ACL 2024long
Processing and reasoning over long contexts is crucial for many practical applications of Large Language Models (LLMs), such as document comprehension and agent construction. Despite recent strides in making LLMs process contexts with more than 100K tokens, there is currently a lack of a standardize…