← Search

Lucrezia Valeriani

2 accepted papers

2025

The Narrow Gate: Localized Image-Text Communication in Native Multimodal Models

NeurIPS 2025poster

Recent advances in multimodal training have significantly improved the integration of image understanding and generation within a unified model. This study investigates how vision-language models (VLMs) handle image-understanding tasks, focusing on how visual information is processed and transferred…

Cited by 0SourceScholar
2023

The geometry of hidden representations of large transformer models

NeurIPS 2023poster

Large transformers are powerful architectures used for self-supervised data analysis across various data types, including protein sequences, images, and text. In these models, the semantic structure of the dataset emerges from a sequence of transformations between one representation and the next. W…

Cited by 51SourcePDFScholar