← Search

Marcel Zalmanovici

1 accepted papers

2025

Exploring Straightforward Methods for Automatic Conversational Red-Teaming

NAACL 2025industry

Large language models (LLMs) are increasingly used in business dialogue systems but they also pose security and ethical risks. Multi-turn conversations, in which context influences the model’s behavior, can be exploited to generate undesired responses. In this paper, we investigate the use of off-th…

Cited by 0SourcePDFScholar