← Search

Saeed Ranjbar Alvar

4 accepted papers

2025

DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models

CVPR 2025poster

Large Multimodal Models (LMMs) have emerged as powerful models capable of understanding various data modalities, including text, images, and videos. LMMs encode both text and visual data into tokens that are then combined and processed by an integrated Large Language Model (LLM). Including visual to…

2024

LaWa: Using Latent Space for In-Generation Image Watermarking

ECCV 2024poster

"With generative models producing high quality images that are indistinguishable from real ones, there is growing concern regarding the malicious usage of AI-generated images. Imperceptible image watermarking is one viable solution towards such concerns. Prior watermarking methods map the image to a…