← Search

George Kour

4 accepted papers

2025

Breaking ReAct Agents: Foot-in-the-Door Attack Will Get You In

NAACL 2025findings

Following the advancement of large language models (LLMs), the development of LLM-based autonomous agents has become prevalent.As a result, the need to understand the security vulnerabilities of these agents has become a critical task. We examine how ReAct agents can be exploited using a straightfor…

Cited by 4SourcePDFScholar
2025

Effective Red-Teaming of Policy-Adherent Agents

EMNLP 2025

Task-oriented LLM-based agents are increasingly used in domains with strict policies, such as refund eligibility or cancellation rules. The challenge lies in ensuring that the agent consistently adheres to these rules and policies, appropriately refusing any request that would violate them, while st

Cited by 0SourcePDFScholar
2025

Exploring Straightforward Methods for Automatic Conversational Red-Teaming

NAACL 2025industry

Large language models (LLMs) are increasingly used in business dialogue systems but they also pose security and ethical risks. Multi-turn conversations, in which context influences the model’s behavior, can be exploited to generate undesired responses. In this paper, we investigate the use of off-th…

Cited by 0SourcePDFScholar
2019

Neural network gradient-based learning of black-box function interfaces

ICLR 2019poster

Deep neural networks work well at approximating complicated functions when provided with data and trained by gradient descent methods. At the same time, there is a vast amount of existing functions that programmatically solve different tasks in a precise manner eliminating the need for training. In…

Cited by 17SourcePDFScholar