← Search

Sebastian Szyller

2 accepted papers

2025

Soft Token Attacks Cannot Reliably Audit Unlearning in Large Language Models

EMNLP 2025

Large language models (LLMs) are trained using massive datasets.However, these datasets often contain undesirable content, e.g., harmful texts, personal information, and copyrighted material.To address this, machine unlearning aims to remove information from trained models.Recent work has shown that

2023

Conflicting Interactions among Protection Mechanisms for Machine Learning Models

AAAI 2023technical

Nowadays, systems based on machine learning (ML) are widely used in different domains. Given their popularity, ML models have become targets for various attacks. As a result, research at the intersection of security/privacy and ML has flourished. Typically such work has focused on individual types o…