← Search

Tarun Gupta

6 accepted papers

2023

Hierarchical Imitation Learning for Stochastic Environments

IROS 2023poster

Many applications of imitation learning require the agent to generate the full distribution of behaviour observed in the training data. For example, to evaluate the safety of autonomous vehicles in simulation, accurate and diverse behaviour models of other road users are paramount. Existing methods…

Cited by 2SourceScholar
2023

Improving Spoken Language Identification with Map-Mix

ICASSP 2023accepted

The pre-trained multi-lingual XLSR model generalizes well for language identification after fine-tuning on unseen languages. However, the performance significantly degrades when the languages are not very distinct from each other, for example, in the case of dialects. Low resource dialect classifica…

Cited by 0SourceScholar
2021

RODE: Learning Roles to Decompose Multi-Agent Tasks

ICLR 2021poster

Role-based learning holds the promise of achieving scalable multi-agent learning by decomposing complex tasks using roles. However, it is largely unclear how to efficiently discover such a set of roles. To solve this problem, we propose to first decompose joint action spaces into restricted role act…

Cited by 260SourcePDFScholar
2021

UneVEn: Universal Value Exploration for Multi-Agent Reinforcement Learning

ICML 2021spotlight

VDN and QMIX are two popular value-based algorithms for cooperative MARL that learn a centralized action value function as a monotonic mixing of per-agent utilities. While this enables easy decentralization of the learned policy, the restricted joint action value function can prevent them from solvi…

Cited by 59SourcePDFScholar