← Search

Paola Merlo

3 accepted papers

2023

BLM-s/lE: A structured dataset of English spray-load verb alternations for testing generalization in LLMs

EMNLP 2023long findings

Current NLP models appear to be achieving performance comparable to human capabilities on well-established benchmarks. New benchmarks are now necessary to test deeper layers of understanding of natural languages by these models. Blackbird's Language Matrices are a recently developed framework th…

Cited by 0SourceScholar
2023

Blackbird language matrices (BLM), a new task for rule-like generalization in neural networks: Can Large Language Models pass the test?

EMNLP 2023long findings

How do we evaluate Large Language Models (LLMs) and determine the aspects and limits of their intelligent behaviour? It is currently conjectured that shortcomings of LLMs in multi-linguality and reasoning are due to a lack of ability to generalize. It has been argued that, instead, humans are bette…

Cited by 0SourceScholar
2021

Multi-Adversarial Learning for Cross-Lingual Word Embeddings

NAACL 2021long

Generative adversarial networks (GANs) have succeeded in inducing cross-lingual word embeddings - maps of matching words across languages - without supervision. Despite these successes, GANs’ performance for the difficult case of distant languages is still not satisfactory. These limitations have be…