← Search

Shurui Li

4 accepted papers

2026

Long Live The Balance: Information Bottleneck Driven Tree-based Policy Optimization

ICML 2026poster

Recent advances in online reinforcement learning (RL) for large language models (LLMs) have demonstrated promising performance in complex reasoning tasks. However, they often exhibit an imbalanced exploration–exploitation trade-off, resulting in unstable optimization and sub-optimal performance. We …

Cited by 0SourceScholar
2026

Recovering Policy-Induced Errors: Benchmarking and Trajectory Synthesis for Robust GUI Agents

ICML 2026spotlight

While GUI agents have advanced rapidly, they often lack the robustness to recover from their own errors, hindering real-world deployment. To bridge this gap at both the evaluation and data levels, we introduce GUI-RobustEval and propose Robustness-driven Trajectory Synthesis. GUI-RobustEval containi…

Cited by 0SourceScholar
2025

Enhanced Data Synthesis for LLM through Reasoning Structures Generated by Hierarchical GFlowNet

ACL 2025finding

Large language models (LLMs) excel in problem-solving but require training data with diverse reasoning processes. Existing methods mainly optimize instruction-response pairs but lack a systematic design for the underlying reasoning structure. This paper proposes RSS: a Reasoning Structure driven dat…

2025

Generative Adversarial Network with Adaptive Synthesis for Brain-Computer Interfaces in Motor Imagery Classification

ICASSP 2025accepted

Motor Imagery (MI) is essential in Brain-Computer Interfaces (BCIs), highlighting the central role of electroencephalography (EEG) in this technology. However, the amount of raw EEG data is often limited. Raw EEG data contains significant noise caused by individual and task-specific differences. The…

Cited by 0SourceScholar