← Search

Yinan Gao

1 accepted papers

2026

Promoting Efficient Reasoning with Verifiable Stepwise Reward

AAAI 2026technical

Large reasoning models (LRMs) have recently achieved significant progress in complex reasoning tasks, aided by reinforcement learning with verifiable rewards. However, LRMs often suffer from overthinking, expending excessive computation on simple problems and reducing efficiency. Existing efficient

Cited by 0SourcePDFScholar