← Search

Johannes Forkel

3 accepted papers

2026

Recurrent Structural Policy Gradient for Partially Observable Mean Field Games

ICML 2026spotlight

Mean Field Games (MFGs) provide a principled framework for modeling interactions in large populations models: at scale, population dynamics become deterministic, with uncertainty entering only through aggregate shocks, or *common noise*. However, algorithmic progress has been limited since model-fre…

Cited by 0SourceScholar
2025

LILO: Learning to Reason at the Frontier of Learnability

NeurIPS 2025poster

Reinforcement learning is widely adopted in post-training large language models, especially for reasoning-style tasks such as maths questions. However, as we show, most existing methods will provably fail to learn from questions that are too hard, where the model always fails, or too easy, where the…

Cited by 0SourceScholar