← Search

Lorenz Hufe

1 accepted papers

2026

Dyslexify: A Mechanistic Defense Against Typographic Attacks in CLIP

ICLR 2026poster

Typographic attacks exploit multi-modal systems by injecting text into images, leading to targeted misclassifications, malicious content generation and even Vision-Language Model jailbreaks. In this work, we analyze how CLIP vision encoders behave under typographic attacks, locating specialized atte…

Cited by 0SourcecodeScholar