← Search

Ruoxue Liu

2 accepted papers

2025

Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

ICLR 2025poster

In offline reinforcement learning, it is necessary to manage out-of-distribution actions to prevent overestimation of value functions. One class of methods, the policy-regularized method, addresses this problem by constraining the target policy to stay close to the behavior policy. Although several…

2024

D2R2: Diffusion-based Representation with Random Distance Matching for Tabular Few-shot Learning

NeurIPS 2024poster

Tabular data is widely utilized in a wide range of real-world applications. The challenge of few-shot learning with tabular data stands as a crucial problem in both industry and academia, due to the high cost or even impossibility of annotating additional samples. However, the inherent heterogeneity…

Cited by 0SourcePDFScholar