2025
Seeing More with Less: Human-like Representations in Vision Models
CVPR 2025highlight
Large multimodal models (LMMs) typically process visual inputs with uniform resolution across the entire field of view, leading to inefficiencies when non-critical image regions are processed as precisely as key areas. Inspired by the human visual system's foveated approach, we apply a sampling meth…