2025
Adaptive Context Length Optimization with Low-Frequency Truncation for Multi-Agent Reinforcement Learning
NeurIPS 2025poster
Recently, deep multi-agent reinforcement learning (MARL) has demonstrated promising performance for solving challenging tasks, such as long-term dependencies and non-Markovian environments. Its success is partly attributed to conditioning policies on large fixed context length. However, such large f…