← Search

Thomas Pouplin

2 accepted papers

2025

The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data

ICML 2025spotlight

Developing autonomous agents capable of performing complex, multi-step decision-making tasks specified in natural language remains a significant challenge, particularly in realistic settings where labeled data is scarce and real-time experimentation is impractical. Existing reinforcement learning (R…

Cited by 0SourcePDFScholar
2024

Relaxed Quantile Regression: Prediction Intervals for Asymmetric Noise

ICML 2024poster

Constructing valid prediction intervals rather than point estimates is a well-established approach for uncertainty quantification in the regression setting. Models equipped with this capacity output an interval of values in which the ground truth target will fall with some prespecified probability.…