← Search

Zhewei Kang

2 accepted papers

2025

Scalable Best-of-N Selection for Large Language Models via Self-Certainty

NeurIPS 2025poster

Best-of-N selection is a key technique for improving the reasoning performance of Large Language Models (LLMs) through increased test-time computation. Current state-of-the-art methods often employ computationally intensive reward models for response evaluation and selection. Reward-free alternative…

Cited by 0SourcecodeScholar