← Search

Eleonora Grassucci

7 accepted papers

2026

Closing the Modality Gap Aligns Group-Wise Semantics

ICLR 2026poster

In multimodal learning, CLIP has been recognized as the \textit{de facto} method for learning a shared latent space across multiple modalities, placing similar representations close to each other and moving them away from dissimilar ones. Although CLIP-based losses effectively align modalities at th…

Cited by 0SourcecodeScholar
2026

TRAINING-FREE MULTIMODAL GUIDANCE FOR VIDEO TO AUDIO GENERATION

ICASSP 2026poster

Video-to-audio (V2A) generation aims to synthesize realistic and semantically aligned audio from silent videos, with potential applications in video editing, Foley sound design, and assistive multimedia. Although the excellent results, existing approaches either require costly joint training on larg…

Cited by 0SourcePDFScholar
2025

A TRIANGLE Enables Multimodal Alignment Beyond Cosine Similarity

NeurIPS 2025poster

Multimodal learning plays a pivotal role in advancing artificial intelligence systems by incorporating information from multiple modalities to build a more comprehensive representation. Despite its importance, current state-of-the-art models still suffer from severe limitations that prevent the succ…

Cited by 0SourceScholar
2025

Gramian Multimodal Representation Learning and Alignment

ICLR 2025poster

Human perception integrates multiple modalities—such as vision, hearing, and language—into a unified understanding of the surrounding reality. While recent multimodal models have achieved significant progress by aligning pairs of modalities via contrastive learning, their solutions are unsuitable wh…

2024

Diffusion Models for Audio Semantic Communication

ICASSP 2024accepted

Directly sending audio signals from a transmitter to a receiver across a noisy channel may absorb consistent bandwidth and be prone to errors when trying to recover the transmitted bits. On the contrary, the recent semantic communication approach proposes to send the semantics and then regenerate se…

Cited by 0SourceScholar
2024

Enhancing Semantic Communication with Deep Generative Models: An Overview

ICASSP 2024accepted

Semantic communication is poised to play a pivotal role in shaping the landscape of future AI-driven communication systems. Its challenge of extracting semantic information from the original complex content and regenerating semantically consistent data at the receiver, possibly being robust to chann…

Cited by 0SourceScholar