2025
Reinforce LLM Reasoning through Multi-Agent Reflection
ICML 2025poster
Leveraging more test-time computation has proven to be an effective way to boost the reasoning capabilities of large language models (LLMs). Among various methods, the verify-and-improve paradigm stands out for enabling dynamic solution exploration and feedback incorporation. However, existing appro…