← Search

Praveen Paruchuri

8 accepted papers

2026

On Discovering Algorithms for Adversarial Imitation Learning

ICLR 2026poster

Adversarial Imitation Learning (AIL) methods, while effective in settings with limited expert demonstrations, are often considered unstable. These approaches typically decompose into two components: Density Ratio (DR) estimation $\frac{\rho_E}{\rho_{\pi}}$, where a discriminator estimates the relati…

Cited by 0SourceScholar
2024

Safety through feedback in Constrained RL

NeurIPS 2024poster

In safety-critical RL settings, the inclusion of an additional cost function is often favoured over the arduous task of modifying the reward function to ensure the agent's safe behaviour. However, designing or evaluating such a cost function can be prohibitively expensive. For instance, in the domai…

2023

City-Scale Pollution Aware Traffic Routing by Sampling Max Flows Using MCMC

AAAI 2023technical

A significant cause of air pollution in urban areas worldwide is the high volume of road traffic. Long-term exposure to severe pollution can cause serious health issues. One approach towards tackling this problem is to design a pollution-aware traffic routing policy that balances multiple objectives…

Cited by 2SourcePDFScholar
2023

Planning and Learning for Non-markovian Negative Side Effects Using Finite State Controllers

AAAI 2023technical

Autonomous systems are often deployed in the open world where it is hard to obtain complete specifications of objectives and constraints. Operating based on an incomplete model can produce negative side effects (NSEs), which affect the safety and reliability of the system. We focus on mitigating NSE…

2022

How Private Is Your RL Policy? An Inverse RL Based Analysis Framework

AAAI 2022technical

Reinforcement Learning (RL) enables agents to learn how to perform various tasks from scratch. In domains like autonomous driving, recommendation systems, and more, optimal RL policies learned could cause a privacy breach if the policies memorize any part of the private reward. We study the set of e…

2022

VidyutVanika21: An Autonomous Intelligent Broker for Smart-grids

IJCAI 2022poster

An autonomous broker that liaises between retail customers and power-generating companies (GenCos) is essential for the smart grid ecosystem. The efficiency brought in by such brokers to the smart grid setup can be studied through a well-developed simulation environment. In this paper, we describe t…

Cited by 7SourcePDFScholar
2021

An Enhanced Advising Model in Teacher-Student Framework using State Categorization

AAAI 2021technical

The teacher-student framework aims to improve the sample efficiency of RL algorithms by deploying an advising mechanism in which a teacher helps a student by guiding its exploration. Prior work in this field has considered an advising mechanism where the teacher advises the student about the optimal…

Cited by 11SourcePDFScholar