FAIRification of LeiLanD
Eric Sanders, Sara Petrollino, Gilles R. Scheifer, Henk van den Heuvel, Christopher Handy
Abstract
LeiLanD (Leiden Language Data) is a searchable catalogue initiated by the Leiden University Centre for Linguistics (LUCL) with the support of CLARIAH. The catalogue contains metadata about language datasets collected at LUCL and other institutes of Leiden University. This paper describes a project to FAIRify the datasets increasing their findability and accessibility through a standardised metadata format CMDI so as to obtain a rich metadata description for all resources and to make them findable through CLARIN’s Virtual Language Observatory. The paper describes the creation of the catalogue and the steps that led from unstructured metadata to CMDI standards. This FAIRifi- cation of LeiLanD has enhanced the findability and accessibility of incredibly diverse collection of language datasets.
BibTeX
@inproceedings{sanders-etal-2024-fairification,
title = "{FAIR}ification of {L}ei{L}an{D}",
author = "Sanders, Eric and
Petrollino, Sara and
Scheifer, Gilles R. and
van den Heuvel, Henk and
Handy, Christopher",
editor = "Calzolari, Nicoletta and
Kan, Min-Yen and
Hoste, Veronique and
Lenci, Alessandro and
Sakti, Sakriani and
Xue, Nianwen",
booktitle = "Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024)",
month = may,
year = "2024",
address = "Torino, Italia",
publisher = "ELRA and ICCL",
url = "https://aclanthology.org/2024.lrec-main.623/",
pages = "7101--7106"
}