← Search

Yingdong Lu

4 accepted papers

2026

Stackelberg Coupling of Online Representation Learning and Reinforcement Learning

ICLR 2026poster

Deep Q-learning jointly learns representations and values within monolithic networks, promising beneficial co-adaptation between features and value estimates. Although this architecture has attained substantial success, the coupling between representation and value learning creates instability as re…

Cited by 0SourceScholar
2025

Optimality and NP-Hardness of Transformers in Learning Markovian Dynamical Functions

NeurIPS 2025poster

Transformer architectures can solve unseen tasks based on input-output pairs in a given prompt due to in-context learning (ICL). Existing theoretical studies on ICL have mainly focused on linear regression tasks, often with i.i.d. inputs. To understand how transformers express in-context learning wh…

Cited by 0SourceScholar