← Search

Jinwang Wu

2 accepted papers

2025

Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation

ICLR 2025poster

First-order logic (FOL) reasoning, which involves sequential deduction, is pivotal for intelligent systems and serves as a valuable task for evaluating reasoning capabilities, particularly in chain-of-thought (CoT) contexts. Existing benchmarks often rely on extensive human annotation or handcrafted…

2023

An Investigation of LLMs’ Inefficacy in Understanding Converse Relations

EMNLP 2023long main

Large Language Models (LLMs) have achieved remarkable success in many formal language oriented tasks, such as structural data-to-text and semantic parsing. However current benchmarks mostly follow the data distribution of the pre-training data of LLMs. Therefore, a natural question rises that do LLM…

Cited by 0SourcecodeScholar