← Search

Clarissa Costen

3 accepted papers

2025

Return Capping: Sample Efficient CVaR Policy Gradient Optimisation

ICML 2025poster

When optimising for conditional value at risk (CVaR) using policy gradients (PG), current methods rely on discarding a large proportion of trajectories, resulting in poor sample efficiency. We propose a reformulation of the CVaR optimisation problem by capping the total return of trajectories used…

2022

Shared Autonomy Systems with Stochastic Operator Models

IJCAI 2022poster

We consider shared autonomy systems where multiple operators (AI and human), can interact with the environment, e.g. by controlling a robot. The decision problem for the shared autonomy system is to select which operator takes control at each timestep, such that a reward specifying the intended syst…

Cited by 14SourcePDFScholar