← Search

Tunga Gungor

3 accepted papers

2025

TR-MTEB: A Comprehensive Benchmark and Embedding Model Suite for Turkish Sentence Representations

EMNLP 2025

We introduce TR-MTEB, the first large-scale, task-diverse benchmark designed to evaluate sentence embedding models for Turkish. Covering six core tasks as classification, clustering, pair classification, retrieval, bitext mining, and semantic textual similarity, TR-MTEB incorporates 26 high-quality

Cited by 0SourcePDFScholar
2024

Evaluating the Quality of a Corpus Annotation Scheme Using Pretrained Language Models

COLING 2024main

Pretrained language models and large language models are increasingly used to assist in a great variety of natural language tasks. In this work, we explore their use in evaluating the quality of alternative corpus annotation schemes. For this purpose, we analyze two alternative annotations of the Tu…

2022

Improving Code-Switching Dependency Parsing with Semi-Supervised Auxiliary Tasks

NAACL 2022findings

Code-switching dependency parsing stands as a challenging task due to both the scarcity of necessary resources and the structural difficulties embedded in code-switched languages. In this study, we introduce novel sequence labeling models to be used as auxiliary tasks for dependency parsing of code-…