2025
VLA-Mark: A cross modal watermark for large vision-language alignment models
EMNLP 2025
Vision-language models demand watermarking solutions that protect intellectual property without compromising multimodal coherence. Existing text watermarking methods disrupt visual-textual alignment through biased token selection and static strategies, leaving semantic-critical concepts vulnerable.