← Search

Daoyi Dong

5 accepted papers

2026

Conditional Diffusion Model for Multi-Agent Dynamic Task Decomposition

AAAI 2026technical

Task decomposition has shown promise in complex cooperative multi-agent reinforcement learning (MARL) tasks, which enables efficient hierarchical learning for long-horizon tasks in dynamic and uncertain environments. However, learning dynamic task decomposition from scratch generally requires a larg

Cited by 0SourcePDFScholar
2025

Mixture-of-Experts Meets In-Context Reinforcement Learning

NeurIPS 2025poster

In-context reinforcement learning (ICRL) has emerged as a promising paradigm for adapting RL agents to downstream tasks through prompt conditioning. However, two notable challenges remain in fully harnessing in-context learning within RL domains: the intrinsic multi-modality of the state-action-rewa…

Cited by 0SourcecodeScholar
2025

PN-GAIL: Leveraging Non-optimal Information from Imperfect Demonstrations

ICLR 2025poster

Imitation learning aims at constructing an optimal policy by emulating expert demonstrations. However, the prevailing approaches in this domain typically presume that the demonstrations are optimal, an assumption that seldom holds true in the complexities of real-world applications. The data collect…

2025

Text-to-Decision Agent: Offline Meta-Reinforcement Learning from Natural Language Supervision

NeurIPS 2025poster

Offline meta-RL usually tackles generalization by inferring task beliefs from high-quality samples or warmup explorations. The restricted form limits their generality and usability since these supervision signals are expensive and even infeasible to acquire in advance for unseen tasks. Learning dire…

Cited by 0SourcecodeScholar