← Search

Narjes Nourzad

2 accepted papers

2026

Memory Based Advantage Shaping for LLM-Guided Reinforcement Learning (Student Abstract)

AAAI 2026technical

In environments with sparse or delayed rewards, reinforcement learning (RL) incurs high sample complexity due to the large number of interactions needed for learning. This limitation has motivated the use of large language models (LLMs) for subgoal discovery and trajectory guidance. While LLMs can s

Cited by 0SourcePDFScholar