← Search

Christian Chiarcos

4 accepted papers

2024

Bridging Computational Lexicography and Corpus Linguistics: A Query Extension for OntoLex-FrAC

COLING 2024main

OntoLex, the dominant community standard for machine-readable lexical resources in the context of RDF, Linked Data and Semantic Web technologies, is currently extended with a designated module for Frequency, Attestations and Corpus-based Information (OntoLex-FrAC). We propose a novel component for O…

Cited by 0SourcePDFScholar
2024

On Modelling Corpus Citations in Computational Lexical Resources

COLING 2024main

In this article we look at how two different standards for lexical resources, TEI and OntoLex, deal with corpus citations in lexicons. We will focus on how corpus citations in retrodigitised dictionaries can be modelled using each of the two standards since this provides us with a suitably challengi…

Cited by 0SourcePDFScholar
2022

Modelling Frequency, Attestation, and Corpus-Based Information with OntoLex-FrAC

COLING 2022main

OntoLex-Lemon has become a de facto standard for lexical resources in the web of data. This paper provides the first overall description of the emerging OntoLex module for Frequency, Attestations, and Corpus-Based Information (OntoLex-FrAC) that is intended to complement OntoLex-Lemon with the neces…

Cited by 21SourcePDFScholar
2020

Towards the First Machine Translation System for Sumerian Transliterations

COLING 2020main

The Sumerian cuneiform script was invented more than 5,000 years ago and represents one of the oldest in history. We present the first attempt to translate Sumerian texts into English automatically. We publicly release high-quality corpora for standardized training and evaluation and report results…