2023
Systematic Assessment of Factual Knowledge in Large Language Models
EMNLP 2023short findings
Previous studies have relied on existing question-answering benchmarks to evaluate the knowledge stored in large language models (LLMs). However, this approach has limitations regarding factual knowledge coverage, as it mostly focuses on generic domains which may overlap with the pretraining data. T…