← Search

Tenglong Liu

4 accepted papers

2025

Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning

IROS 2025

In offline reinforcement learning, value overestimation caused by out-of-distribution (OOD) actions significantly limits policy performance. Recently, diffusion models have been leveraged for their strong distribution-matching capabilities, enforcing conservatism through behavior policy constraints.

Cited by 0SourceScholar
2025

Learning Predictive Control with Online Modeling for Agile Maneuvering of Autonomous Vehicles

IROS 2025

The agile maneuvering control of autonomous vehicles (AVs) requires the tracking of reference trajectories characterized by high acceleration, sharp curvature, considerable disturbances, and significant time-varying, all while ensuring stability and accuracy. The inherent uncertainty and time-varyin

Cited by 0SourceScholar
2025

Skill Expansion and Composition in Parameter Space

ICLR 2025poster

Humans excel at reusing prior knowledge to address new challenges and developing skills while solving problems. This paradigm becomes increasingly popular in the development of autonomous agents, as it develops systems that can self-evolve in response to new challenges like human beings. However, pr…

2024

Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning

ICML 2024poster

In offline reinforcement learning, the challenge of out-of-distribution (OOD) is pronounced. To address this, existing methods often constrain the learned policy through policy regularization. However, these methods often suffer from the issue of unnecessary conservativeness, hampering policy improv…