NAACL 2024findings6 citations

Graph-Induced Syntactic-Semantic Spaces in Transformer-Based Variational AutoEncoders

Yingji Zhang, Marco Valentino, Danilo Carvalho, Ian Pratt-Hartmann, Andre Freitas

Abstract

The injection of syntactic information in Variational AutoEncoders (VAEs) can result in an overall improvement of performances and generalisation. An effective strategy to achieve such a goal is to separate the encoding of distributional semantic features and syntactic structures into heterogeneous latent spaces via multi-task learning or dual encoder architectures. However, existing works employing such techniques are limited to LSTM-based VAEs. This work investigates latent space separation methods for structural syntactic injection in Transformer-based VAE architectures (i.e., Optimus) through the integration of graph-based models. Our empirical evaluation reveals that the proposed end-to-end VAE architecture can improve theoverall organisation of the latent space, alleviating the information loss occurring in standard VAE setups, and resulting in enhanced performances on language modelling and downstream generation tasks.

BibTeX
@inproceedings{zhang-etal-2024-graph,
    title = "Graph-Induced Syntactic-Semantic Spaces in Transformer-Based Variational {A}uto{E}ncoders",
    author = "Zhang, Yingji  and
      Valentino, Marco  and
      Carvalho, Danilo  and
      Pratt-Hartmann, Ian  and
      Freitas, Andre",
    editor = "Duh, Kevin  and
      Gomez, Helena  and
      Bethard, Steven",
    booktitle = "Findings of the Association for Computational Linguistics: NAACL 2024",
    month = jun,
    year = "2024",
    address = "Mexico City, Mexico",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2024.findings-naacl.32/",
    doi = "10.18653/v1/2024.findings-naacl.32",
    pages = "474--489"
}
Graph-Induced Syntactic-Semantic Spaces in Transformer-Based Variational AutoEncoders · NAACL 2024