← Search

Marian Simko

5 accepted papers

2025

Large Language Models for Multilingual Previously Fact-Checked Claim Detection

EMNLP 2025

In our era of widespread false information, human fact-checkers often face the challenge of duplicating efforts when verifying claims that may have already been addressed in other countries or languages. As false information transcends linguistic boundaries, the ability to automatically detect previ

2025

skLEP: A Slovak General Language Understanding Benchmark

ACL 2025finding

In this work, we introduce skLEP, the first comprehensive benchmark specifically designed for evaluating Slovak natural language understanding (NLU) models. We have compiled skLEP to encompass nine diverse tasks that span token-level, sentence-pair, and document-level challenges, thereby offering a…

2024

Women Are Beautiful, Men Are Leaders: Gender Stereotypes in Machine Translation and Language Modeling

EMNLP 2024finding

We present GEST – a new manually created dataset designed to measure gender-stereotypical reasoning in language models and machine translation systems. GEST contains samples for 16 gender stereotypes about men and women (e.g., Women are beautiful, Men are leaders) that are compatible with the Englis…

2022

SlovakBERT: Slovak Masked Language Model

EMNLP 2022finding

We introduce a new Slovak masked language model called SlovakBERT. This is to our best knowledge the first paper discussing Slovak transformers-based language models. We evaluate our model on several NLP tasks and achieve state-of-the-art results. This evaluation is likewise the first attempt to est…