2025
Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study
EMNLP 2025
Logical reasoning is a core capability for large language models (LLMs), yet existing benchmarks that rely solely on final-answer accuracy fail to capture the quality of the reasoning process. To address this, we introduce FineLogic, a fine-grained evaluation framework that assesses logical reasonin