← Search

Zhuojian Li

2 accepted papers

2026

Multi-timescale Reinforcement Learning by Value Reconstruction

ICML 2026poster

Most reinforcement learning (RL) baselines maximize future cumulative rewards with a fixed single discount factor, which limits their performance in complex sequential decision-making tasks due to a failure to balance short-term objectives and long-term planning. To address this issue, this paper fo…

Cited by 0SourceScholar
2026

Structured Expert Routing with Multi-View Task Priors for Offline Meta-Reinforcement Learning

ICML 2026poster

Offline meta-reinforcement learning requires agents to generalize to unseen tasks from fixed datasets, yet existing sequence-based and MoE-based methods rely on implicit or token-level routing signals that fail to capture task-level structure. We propose the **Task-Guided Router (TGR)**, a structure…

Cited by 0SourceScholar