← Search

Ruishuo Chen

3 accepted papers

2026

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

ICML 2026poster

Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward queries are infeasible, these models must be trained using static offline datasets. Prevailing training methods typically rely on a proxy model to provide reward fe…

Cited by 0SourceScholar
2026

PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching

ICML 2026poster

Unsupervised Reinforcement Learning from Internal Feedback (RLIF) has emerged as a promising paradigm for eliciting the latent capabilities of Large Language Models (LLMs) without external supervision. However, current methods rely on heuristic intrinsic rewards, which often lack a well-defined theo…

Cited by 0SourceScholar
2024

Provably and Practically Efficient Adversarial Imitation Learning with General Function Approximation

NeurIPS 2024poster

As a prominent category of imitation learning methods, adversarial imitation learning (AIL) has garnered significant practical success powered by neural network approximation. However, existing theoretical studies on AIL are primarily limited to simplified scenarios such as tabular and linear functi…

Cited by 1SourcePDFScholar