← Search

Grzegorz Kondrak

8 accepted papers

2025

Semi-Automated Construction of Sense-Annotated Datasets for Practically Any Language

COLING 2025main

High-quality sense-annotated datasets are vital for evaluating and comparing WSD systems. We present a novel approach to creating parallel sense-annotated datasets, which can be applied to any language that English can be translated into. The method incorporates machine translation, word alignment,…

2024

Semantically-Prompted Language Models Improve Visual Descriptions

NAACL 2024findings

Language-vision models like CLIP have made significant strides in vision tasks, such as zero-shot image classification (ZSIC). However, generating specific and expressive visual descriptions remains challenging; descriptions produced by current methods are often ambiguous and lacking in granularity.…

Cited by 0SourcePDFScholar
2024

Translation-based Lexicalization Generation and Lexical Gap Detection: Application to Kinship Terms

ACL 2024long

Constructing lexicons with explicitly identified lexical gaps is a vital part of building multilingual lexical resources. Prior work has leveraged bilingual dictionaries and linguistic typologies for semi-automatic identification of lexical gaps. Instead, we propose a generally-applicable algorithmi…

Cited by 1SourcePDFScholar
2023

Don’t Trust ChatGPT when your Question is not in English: A Study of Multilingual Abilities and Types of LLMs

EMNLP 2023long main

Large language models (LLMs) have demonstrated exceptional natural language understanding abilities, and have excelled in a variety of natural language processing (NLP) tasks. Despite the fact that most LLMs are trained predominantly on English, multiple studies have demonstrated their capabilities…

Cited by 0SourceScholar
2022

Improving HowNet-Based Chinese Word Sense Disambiguation with Translations

EMNLP 2022finding

Word sense disambiguation (WSD) is the task of identifying the intended sense of a word in context. While prior work on unsupervised WSD has leveraged lexical knowledge bases, such as WordNet and BabelNet, these resources have proven to be less effective for Chinese. Instead, the most widely used le…

Cited by 10SourcePDFScholar