2026
Geometry-Preserving Unsupervised Alignment for Heterogeneous Foundation Models
ICML 2026poster
Foundation models have driven rapid progress in computer vision, yet the two dominant paradigm, vision-language foundation models (VLMs) and vision-only foundation models (VFMs), remain only partially compatible. VLMs offer language-grounded semantic alignment but are often visually coarse, while VF…