2024
Offline Primal-Dual Reinforcement Learning for Linear MDPs
AISTATS 2024poster
Offline Reinforcement Learning (RL) aims to learn a near-optimal policy from a fixed dataset of transitions collected by another policy. This problem has attracted a lot of attention recently, but most existing methods with strong theoretical guarantees are restricted to finite-horizon or tabular se…