← Search

Krisztian Flautner

2 accepted papers

2025

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

ACL 2025long

Evaluating large language model (LLM) is a complex task. Pairwise ranking has emerged as state-of-the-art method to evaluate human preferences by having humans compare pairs of LLM outputs based on predefined criteria, enabling ranking across multiple LLMs by aggregating pairwise results through alg…

2023

Label Agnostic Pre-training for Zero-shot Text Classification

ACL 2023findings

Conventional approaches to text classification typically assume the existence of a fixed set of predefined labels to which a given text can be classified. However, in real-world applications, there exists an infinite label space for describing a given text. In addition, depending on the aspect (sent…