Faithful Self-Refinement in Mathematical Reasoning via Progressive Back-Translation
Large language models (LLMs) can achieve superior results through iterative refinement based on internal or external signals, compared to the unstable outputs from a single pass. However, the reliability of existing internal signals is questionable due to their susceptibility to intrinsic hallucinat…