2025
Strategist: Self-improvement of LLM Decision Making via Bi-Level Tree Search
ICLR 2025poster
Traditional reinforcement learning and planning require a lot of data and training to develop effective strategies. On the other hand, large language models (LLMs) can generalize well and perform tasks without prior training but struggle with complex planning and decision-making. We introduce **STRA…