← Search

Yanxiao Zhao

1 accepted papers

2026

ComputerRL: Scaling End-to-End Online Reinforcement Learning for Computer Use Agents

ICLR 2026poster

We introduce ComputerRL, a framework for autonomous desktop intelligence that enables agents to operate complex digital workspaces skillfully. ComputerRL features the API-GUI paradigm, which unifies programmatic API calls and direct GUI interaction to address the inherent mismatch between machine ag…

Cited by 0SourcecodeScholar