2025
SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
NeurIPS 2025poster
The recent DeepSeek-R1 release has demonstrated the immense potential of reinforcement learning (RL) in enhancing the general reasoning capabilities of large language models (LLMs). While DeepSeek-R1 and other follow-up work primarily focus on applying RL to competitive coding and math problems, thi…