Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis
Robust reinforcement learning (RL) under the average-reward criterion is essential for long-term decision-making, particularly when the environment may differ from its training dynamics. However, most existing studies focus on model-based settings and provide only asymptotic guarantees, hindering th…