← Search

ZihaoLian

1 accepted papers

2025

Test-time Adapted Reinforcement Learning with Action Entropy Regularization

ICML 2025poster

Offline reinforcement learning is widely applied in multiple fields due to its advantages in efficiency and risk control. However, a major problem it faces is the distribution shift between offline datasets and online environments. This mismatch leads to out-of-distribution (OOD) state-action pairs…

Cited by 0SourcePDFScholar