2023
WikiWhy: Answering and Explaining Cause-and-Effect Questions
ICLR 2023top-5%
As large language models (LLMs) grow larger and more sophisticated, assessing their "reasoning" capabilities in natural language grows more challenging. Recent question answering (QA) benchmarks that attempt to assess reasoning are often limited by a narrow scope of covered situations and subject ma…