← Search

Qirun Dai

3 accepted papers

2026

LLMs Must Think Thrice to Solve Executable Counterfactuals

ICLR 2026poster

Counterfactual reasoning, a hallmark of intelligence, consists of three steps: inferring latent variables from observations (abduction), constructing alternative situations (interventions), and predicting the outcomes of the alternatives (prediction). This skill is essential for advancing LLMs' caus…

Cited by 0SourceScholar
2025

Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities

EMNLP 2025

Selecting appropriate training data is crucial for instruction fine-tuning of large language models (LLMs), which aims to (1) elicit strong capabilities, and (2) achieve balanced performance across different tasks. Influence-based methods show promise in achieving (1), by estimating the contribution

Cited by 0SourcePDFScholar