← Search

Shivam Singhal

5 accepted papers

2025

Correlated Proxies: A New Definition and Improved Mitigation for Reward Hacking

ICLR 2025spotlight

Because it is difficult to precisely specify complex objectives, reinforcement learning policies are often optimized using proxy reward functions that only approximate the true goal. However, optimizing proxy rewards frequently leads to reward hacking: the optimized reward function ceases to be a go…

2023

A Patient Invariant Model Towards the Prediction of Freezing of Gait

ICASSP 2023accepted

Freezing of Gait (FoG) is one of the incapacitating motor symptoms that appear in patients with Parkinson’s Disease (PD). FoG manifests gait impairments and imposes unforeseen difficulties in commencing the locomotion. Frequent episodes of FoG often lead to fall-related injuries and impart dreadful…

Cited by 0SourceScholar
2021

A Patient-Invariant Model for Freezing of Gait Detection Aided by Wavelet Decomposition

ICASSP 2021accepted

Freezing of Gait (FoG) is a paroxysmal and devitalizing symptom associated with Parkinson’s disease (PD). Episodes of FoG impedes gait and augments fall propensity, often leading to serious fall-injury. In this paper, we present a method for online detection of FoG using a wearable motion sensor. Th…

Cited by 0SourceScholar
2019

Learning Deep Visuomotor Policies for Dexterous Hand Manipulation

ICRA 2019poster

Multi-fingered dexterous hands are versatile and capable of acquiring a diverse set of skills such as grasping, in-hand manipulation, and tool use. To fully utilize their versatility in real-world scenarios, we require algorithms and policies that can control them using on-board sensing capabilities…

Cited by 62SourceScholar