← Search

Arnav Jhala

1 accepted papers

2025

Action-Dependent Optimality-Preserving Reward Shaping

ICML 2025poster

Recent RL research has utilized reward shaping--particularly complex shaping rewards such as intrinsic motivation (IM)--to encourage agent exploration in sparse-reward environments. While often effective, ``reward hacking'' can lead to the shaping reward being optimized at the expense of the extrins…

Cited by 0SourcePDFScholar