← Search

David Roberts

2 accepted papers

2025

Action-Dependent Optimality-Preserving Reward Shaping

ICML 2025poster

Recent RL research has utilized reward shaping--particularly complex shaping rewards such as intrinsic motivation (IM)--to encourage agent exploration in sparse-reward environments. While often effective, ``reward hacking'' can lead to the shaping reward being optimized at the expense of the extrins…

Cited by 0SourcePDFScholar
2015

CoCE-SMART: Consensus clustering based on enhanced splitting-merging awareness tactics

ICASSP 2015accepted

In this paper, we propose a new consensus clustering algorithm, which is based on an existing clustering paradigm, called enhanced splitting merging awareness tactics (E-SMART). The problem of determining the number of clusters, which affects many state-of-theart consensus clustering algorithms, is…

Cited by 0SourceScholar