← Search

Ramya Keerthy Thatikonda

2 accepted papers

2025

Logical Reasoning with Outcome Reward Models for Test-Time Scaling

EMNLP 2025

Logical reasoning is a critical benchmark for evaluating the capabilities of large language models (LLMs), as it reflects their ability to derive valid conclusions from given premises. While the combination of test-time scaling with dedicated outcome or process reward models has opened up new avenue

Cited by 0SourcePDFScholar