← Search

Itay Nakash

2 accepted papers

2025

Breaking ReAct Agents: Foot-in-the-Door Attack Will Get You In

NAACL 2025findings

Following the advancement of large language models (LLMs), the development of LLM-based autonomous agents has become prevalent.As a result, the need to understand the security vulnerabilities of these agents has become a critical task. We examine how ReAct agents can be exploited using a straightfor…

Cited by 4SourcePDFScholar
2025

Effective Red-Teaming of Policy-Adherent Agents

EMNLP 2025

Task-oriented LLM-based agents are increasingly used in domains with strict policies, such as refund eligibility or cancellation rules. The challenge lies in ensuring that the agent consistently adheres to these rules and policies, appropriately refusing any request that would violate them, while st

Cited by 0SourcePDFScholar