2026
Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization
ICLR 2026poster
Vision–Language Models (VLMs) have become essential for tasks such as image synthesis, captioning, and retrieval by aligning textual and visual information in a shared embedding space. Yet, this flexibility also makes them vulnerable to malicious prompts designed to produce unsafe content, raising c…