← Search

Sumeet Ramesh Motwani

4 accepted papers

2025

REAL: Benchmarking Autonomous Agents on Deterministic Simulations of Real Websites

NeurIPS 2025poster

We introduce REAL, a benchmark and framework for multi-turn agent evaluations on deterministic simulations of real-world websites. REAL comprises high-fidelity, deterministic replicas of 11 widely-used websites across domains such as e-commerce, travel, communication, and professional networking. We…

Cited by 0SourceScholar
2024

STARC: A General Framework For Quantifying Differences Between Reward Functions

ICLR 2024poster

In order to solve a task using reinforcement learning, it is necessary to first formalise the goal of that task as a *reward function*. However, for many real-world tasks, it is very difficult to manually specify a reward function that never incentivises undesirable behaviour. As a result, it is inc…

Cited by 9SourcePDFScholar
2024

Secret Collusion among AI Agents: Multi-Agent Deception via Steganography

NeurIPS 2024poster

Recent advancements in generative AI suggest the potential for large-scale interaction between autonomous agents and humans across platforms such as the internet. While such interactions could foster productive cooperation, the ability of AI agents to circumvent security oversight raises critical mu…

Cited by 6SourcePDFScholar
2024

Unelicitable Backdoors via Cryptographic Transformer Circuits

NeurIPS 2024poster

The rapid proliferation of open-source language models significantly increases the risks of downstream backdoor attacks. These backdoors can introduce dangerous behaviours during model deployment and can evade detection by conventional cybersecurity monitoring systems. In this paper, we introduce a…

Cited by 5SourcePDFScholar