← Search

Sanket Gaurav

3 accepted papers

2025

Prompted Policy Search: Reinforcement Learning through Linguistic and Numerical Reasoning in LLMs

NeurIPS 2025poster

Reinforcement Learning (RL) traditionally relies on scalar reward signals, limiting its ability to leverage the rich semantic knowledge often available in real-world tasks. In contrast, humans learn efficiently by combining numerical feedback with language, prior knowledge, and common sense. We intr…

Cited by 0SourceScholar
2023

Robot Learning to Mop Like Humans Using Video Demonstrations

IROS 2023poster

Though mopping the floor is a mundane and tedious daily task, enabling robots to perform it comparably to humans remains a challenge. Hand-coding desired mopping behaviors for variable surfaces and situations is particularly difficult. In this paper, we develop a robotic system for mopping the floor…

Cited by 1SourceScholar
2017

Goal-predictive robotic teleoperation from noisy sensors

ICRA 2017poster

Robotic teleoperation from a human operator's pose demonstrations provides an intuitive and effective means of control that has been made feasible by improvements in sensor technologies in recent years. However, the imprecision of low-cost depth cameras and the difficulty of calibrating a frame of r…

Cited by 27SourceScholar