← Search

Piek Vossen

4 accepted papers

2025

Language Models Lack Temporal Generalization and Bigger is Not Better

ACL 2025finding

This paper presents elaborate testing of various LLMs on their generalization capacities. We finetune six encoder models that have been pretrained with very different data (varying in size, language, and period) on a challenging event detection task in Early Modern Dutch archival texts. Each model i…

2023

A Machine with Short-Term, Episodic, and Semantic Memory Systems

AAAI 2023technical

Inspired by the cognitive science theory of the explicit human memory systems, we have modeled an agent with short-term, episodic, and semantic memory systems, each of which is modeled with a knowledge graph. To evaluate this system and analyze the behavior of this agent, we designed and released ou…

2023

Reasoning about Ambiguous Definite Descriptions

EMNLP 2023short findings

Natural language reasoning plays an increasingly important role in improving language models' ability to solve complex language understanding tasks. An interesting use case for reasoning is the resolution of context-dependent ambiguity. But no resources exist to evaluate how well Large Language Mode…

Cited by 0SourcecodeScholar
2020

Would you describe a leopard as yellow? Evaluating crowd-annotations with justified and informative disagreement

COLING 2020main

Semantic annotation tasks contain ambiguity and vagueness and require varying degrees of world knowledge. Disagreement is an important indication of these phenomena. Most traditional evaluation methods, however, critically hinge upon the notion of inter-annotator agreement. While alternative framewo…