← Search

Anirban Santara

2 accepted papers

2023

A Contextual Bandit Approach for Learning to Plan in Environments with Probabilistic Goal Configurations

ICRA 2023poster

Object-goal navigation (Object-nav) entails searching, recognizing and navigating to a target object. Object-nav has been extensively studied by the Embodied-AI community, but most solutions are often restricted to considering static objects (e.g., television, fridge, etc.), We propose a modular fra…

Cited by 7SourceScholar
2022

Learning to Plan Variable Length Sequences of Actions with a Cascading Bandit Click Model of User Feedback

AISTATS 2022poster

Motivated by problems of ranking with partial information, we introduce a variant of the cascading bandit model that considers flexible length sequences with varying rewards and losses. We formulate two generative models for this problem within the generalized linear setting, and design and analyze…

Cited by 4SourcePDFScholar