← Search

Kenneth Styppa

1 accepted papers

2026

Process Reward Agents for Steering Knowledge-Intensive Reasoning

ICML 2026poster

Reasoning in knowledge-intensive domains remains challenging because intermediate steps are often not locally verifiable: unlike math or code, evaluating step correctness may require synthesizing clues across large external knowledge sources. As a result, subtle errors can propagate through reasonin…

Cited by 0SourceScholar