← Search

Thommen Karimpanal George

4 accepted papers

2024

EMOTE: An Explainable Architecture for Modelling the Other through Empathy

IJCAI 2024poster

Empathy allows us to assume others are like us and have goals analogous to our own. This can also at times be applied to multi-agent games - e.g. Agent 1's attraction to green balls is analogous to Agent 2's attraction to red balls. Drawing inspiration from empathy, we propose EMOTE, a simple and…

2022

Learning to Constrain Policy Optimization with Virtual Trust Region

NeurIPS 2022accept

We introduce a constrained optimization method for policy gradient reinforcement learning, which uses two trust regions to regulate each policy update. In addition to using the proximity of one single old policy as the first trust region as done by prior works, we propose forming a second trust regi…

Cited by 5SourcePDFScholar
2021

A New Representation of Successor Features for Transfer across Dissimilar Environments

ICML 2021spotlight

Transfer in reinforcement learning is usually achieved through generalisation across tasks. Whilst many studies have investigated transferring knowledge when the reward function changes, they have assumed that the dynamics of the environments remain consistent. Many real-world RL problems require tr…

Cited by 23SourcePDFScholar
2021

Model-Based Episodic Memory Induces Dynamic Hybrid Controls

NeurIPS 2021poster

Episodic control enables sample efficiency in reinforcement learning by recalling past experiences from an episodic memory. We propose a new model-based episodic memory of trajectories addressing current limitations of episodic control. Our memory estimates trajectory values, guiding the agent towar…

Cited by 21SourcePDFScholar