← Search

Savvas Zannettou

2 accepted papers

2025

Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification

EMNLP 2025

Recently, autonomous agents built on large language models (LLMs) have experienced significant development and are being deployed in real-world applications. Through the usage of tools, these systems can perform actions in the real world. Given the agents’ practical applications and ability to execu

2025

Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions

ICCV 2025poster

Recent advances in text-to-image diffusion models have enabled the creation of a new form of digital art: optical illusions---visual tricks that create different perceptions of reality. However, adversaries may misuse such techniques to generate hateful illusions, which embed specific hate messages…

Cited by 0SourcePDFScholar