← Search

Rishabh Agrawal

6 accepted papers

2026

Towards Offline Imitation Learning: Strictly Batch Settings, Generalization, and Variability in Expertise

AAAI 2026technical

My doctoral research develops a unified framework for offline imitation learning (IL) that tackles three central challenges: achieving sample efficiency in strictly batch settings, ensuring robustness and generalization under dynamics shifts, and learning from demonstrations of varying quality. At t

Cited by 0SourcePDFScholar
2025

Markov Balance Satisfaction Improves Performance in Strictly Batch Offline Imitation Learning

AAAI 2025technical

Imitation learning (IL) is notably effective for robotic tasks where directly programming behaviors or defining optimal control costs is challenging. In this work, we address a scenario where the imitator relies solely on observed behavior and cannot make environmental interactions during learning.…

2025

SpecMAS: A Multi-Agent System for Self-Verifying System Generation via Formal Model Checking

NeurIPS 2025poster

We present SpecMAS, a novel multi-agent system that autonomously constructs and formally verifies executable system models from natural language specifications. Given a Standard Operating Procedure (SOP) describing a target system, SpecMAS parses the specification, identifies relevant operational mo…

Cited by 0SourceScholar
2024

PFA-ERC: Psuedo-Future Augmented Dynamic Emotion Recognition in Conversations

EMNLP 2024finding

AI systems’ ability to interpret human emotions and adapt to variations is becoming more crucial as AI gets embedded into everyone’s daily lives. Emotion Recognition in Conversations (ERC) is based on this fundamental challenge. Current state-of-the-art technologies in ERC are limited due to the nee…

Cited by 0SourcePDFScholar
2022

Socially Intelligent Genetic Agents for the Emergence of Explicit Norms

IJCAI 2022poster

Norms help regulate a society. Norms may be explicit (represented in structured form) or implicit. We address the emergence of explicit norms by developing agents who provide and reason about explanations for norm violations in deciding sanctions and identifying alternative norms. These agents use…