← Search

Yongkang Du

1 accepted papers

2024

Self-contradictory reasoning evaluation and detection

EMNLP 2024finding

In a plethora of recent work, large language models (LLMs) demonstrated impressive reasoning ability, but many proposed downstream reasoning tasks only focus on performance-wise evaluation. Two fundamental questions persist: 1) how consistent is the reasoning, and 2) can models detect unreliable rea…