PQDA:Policy-Aligned Q-Consistency Meets Decoupled Augmentation for Generalizable Visual RL
A fundamental challenge in visual reinforcement learning (RL) is achieving robust generalization across environments with varying visual distractions. Current RL methods struggle with generalization due to their inability to differentiate foreground and background features during augmentation,while