Breaking the Computational Barrier: Provably Efficient Actor–Critic for Low-Rank MDPs
Reinforcement learning (RL) is a fundamental framework for sequential decision-making, in which an agent learns an optimal policy through interactions with an unknown environment. In settings with function approximation, many existing RL algorithms achieve favorable sample complexity, but often rely…