← Search

Martin Corredor

1 accepted papers

2023

ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning

ICLR 2023top-25%

Large language models show improved downstream task performance when prompted to generate step-by-step reasoning to justify their final answers. These reasoning steps greatly improve model interpretability and verification, but objectively studying their correctness (independent of the final answer)…