← Search

Julien Roy

5 accepted papers

2025

Efficient Biological Data Acquisition through Inference Set Design

ICLR 2025poster

In drug discovery, highly automated high-throughput laboratories are used to screen a large number of compounds in search of effective drugs. These experiments are expensive, so one might hope to reduce their cost by only experimenting on a subset of the compounds, and predicting the outcomes of the…

Cited by 0SourcePDFScholar
2025

SynFlowNet: Design of Diverse and Novel Molecules with Synthesis Constraints

ICLR 2025spotlight

Generative models see increasing use in computer-aided drug design. However, while performing well at capturing distributions of molecular motifs, they often produce synthetically inaccessible molecules. To address this, we introduce SynFlowNet, a GFlowNet model whose action space uses chemical reac…

2022

Direct Behavior Specification via Constrained Reinforcement Learning

ICML 2022spotlight

The standard formulation of Reinforcement Learning lacks a practical way of specifying what are admissible and forbidden behaviors. Most often, practitioners go about the task of behavior specification by manually engineering the reward function, a counter-intuitive process that requires several ite…

2020

Adversarial Soft Advantage Fitting: Imitation Learning without Policy Optimization

NeurIPS 2020spotlight

Adversarial Imitation Learning alternates between learning a discriminator -- which tells apart expert's demonstrations from generated ones -- and a generator's policy to produce trajectories that can fool this discriminator. This alternated optimization is known to be delicate in practice since it…

2020

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning

NeurIPS 2020poster

In multi-agent reinforcement learning, discovering successful collective behaviors is challenging as it requires exploring a joint action space that grows exponentially with the number of agents. While the tractability of independent agent-wise exploration is appealing, this approach fails on tasks…

Cited by 31SourcePDFScholar