← Search

Anders S{\o}gaard

4 accepted papers

2025

Lost in Embeddings: Information Loss in Vision–Language Models

EMNLP 2025

Vision–language models (VLMs) often process visual inputs through a pretrained vision encoder, followed by a projection into the language model’s embedding space via a connector component. While crucial for modality fusion, the potential information loss induced by this projection step and its direc

2025

What if Othello-Playing Language Models Could See?

EMNLP 2025

Language models are often said to face a symbol grounding problem. While some have argued the problem can be solved without resort to other modalities, many have speculated that grounded learning is more efficient. We explore this question in Othello, a simplified, rule-based world that offers a con