← Search

Björn Haddenhorst

5 accepted papers

2024

Identifying Copeland Winners in Dueling Bandits with Indifferences

AISTATS 2024poster

We consider the task of identifying the Copeland winner(s) in a dueling bandits problem with ternary feedback. This is an underexplored but practically relevant variant of the conventional dueling bandits problem, in which, in addition to strict preference between two arms, one may observe feedback…

Cited by 2SourcePDFScholar
2023

AC-Band: A Combinatorial Bandit-Based Approach to Algorithm Configuration

AAAI 2023technical

We study the algorithm configuration (AC) problem, in which one seeks to find an optimal parameter configuration of a given target algorithm in an automated way. Although this field of research has experienced much progress recently regarding approaches satisfying strong theoretical guarantees, ther…

2022

Finding Optimal Arms in Non-stochastic Combinatorial Bandits with Semi-bandit Feedback and Finite Budget

NeurIPS 2022accept

We consider the combinatorial bandits problem with semi-bandit feedback under finite sampling budget constraints, in which the learner can carry out its action only for a limited number of times specified by an overall budget. The action is to choose a set of arms, whereupon feedback for each arm in…

Cited by 14SourcePDFScholar
2021

Identification of the Generalized Condorcet Winner in Multi-dueling Bandits

NeurIPS 2021poster

The reliable identification of the “best” arm while keeping the sample complexity as low as possible is a common task in the field of multi-armed bandits. In the multi-dueling variant of multi-armed bandits, where feedback is provided in the form of a winning arm among as set of k chosen ones, a rea…

2021

Testification of Condorcet Winners in dueling bandits

UAI 2021poster

Several algorithms for finding the best arm in the dueling bandits setting assume the existence of a Condorcet winner (CW), that is, an arm that uniformly dominates all other arms. Yet, by simply relying on this assumption but not verifying it, such algorithms may produce doubtful results in cases w…

Cited by 7SourcePDFScholar