← Search

Jonathan Martinez

3 accepted papers

2026

Decision Transformers As Zero-Shot Learners via Text-Behavior Alignment

ICML 2026spotlight

Offline meta-reinforcement learning (meta-RL) aims to train agents that can generalize to unseen tasks using pre-collected data from related tasks. Recent approaches leverage the scalability of transformer architectures to model behavior sequences and support task adaptation using target task demons…

Cited by 0SourceScholar
2021

Improving the Performance-Compatibility Tradeoff with Personalized Objective Functions

AAAI 2021technical

AI-systems that model and interact with their users can up-date their models over time to reflect new information and changes in the environment. Although these updates may improve the overall performance of the AI-system, they may actually hurt the performance with respect to individual users. Prio…

Cited by 5SourcePDFScholar