2023
This is not a Dataset: A Large Negation Benchmark to Challenge Large Language Models
EMNLP 2023long main
Although large language models (LLMs) have apparently acquired a certain level of grammatical knowledge and the ability to make generalizations, they fail to interpret negation, a crucial step in Natural Language Processing. We try to clarify the reasons for the sub-optimal performance of LLMs under…