← Search

Jayaraman Dinesh

1 accepted papers

2022

Conservative and Adaptive Penalty for Model-Based Safe Reinforcement Learning

AAAI 2022technical

Reinforcement Learning (RL) agents in the real world must satisfy safety constraints in addition to maximizing a reward objective. Model-based RL algorithms hold promise for reducing unsafe real-world actions: they may synthesize policies that obey all constraints using simulated samples from a lear…