2026
SWE-MiniSandbox: Container-Free Reinforcement Learning for Building Software Engineering Agents
ICML 2026poster
Reinforcement learning (RL) has become a key paradigm for training software engineering (SWE) agents, yet its practical accessibility and scalability is often constrained by container-based execution frameworks used for environment isolation. As the number of task instances increases, pre-cached con…