2026
MIRA: Memory-Integrated Reinforcement Learning Agent with Limited LLM Guidance
ICLR 2026poster
Reinforcement learning (RL) agents often face high sample complexity in sparse or delayed reward settings, due to limited prior knowledge. Conversely, large language models (LLMs) can provide subgoal structures, plausible trajectories, and abstract priors that support early learning. Yet heavy relia…