← Search

Shukuan Wang

2 accepted papers

2026

GIPO: Gaussian Importance Sampling Policy Optimization

ICML 2026poster

Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation. However, RL remains limited by poor data efficiency, particularly in settings where interaction data are scarce and quickly become outdated. To address this c…

Cited by 0SourceScholar
2024

Monte Carlo Tree Search based Space Transfer for Black Box Optimization

NeurIPS 2024spotlight

Bayesian optimization (BO) is a popular method for computationally expensive black-box optimization. However, traditional BO methods need to solve new problems from scratch, leading to slow convergence. Recent studies try to extend BO to a transfer learning setup to speed up the optimization, where…