← Search

Till J. Bungert

4 accepted papers

2024

Navigating the Maze of Explainable AI: A Systematic Approach to Evaluating Methods and Metrics

NeurIPS 2024poster

Explainable AI (XAI) is a rapidly growing domain with a myriad of proposed methods as well as metrics aiming to evaluate their efficacy. However, current studies are often of limited scope, examining only a handful of XAI methods and ignoring underlying design parameters for performance, such as the…

2024

Overcoming Common Flaws in the Evaluation of Selective Classification Systems

NeurIPS 2024spotlight

Selective Classification, wherein models can reject low-confidence predictions, promises reliable translation of machine-learning based classification systems to real-world scenarios such as clinical diagnostics. While current evaluation of these systems typically assumes fixed working points based…

2023

A Call to Reflect on Evaluation Practices for Failure Detection in Image Classification

ICLR 2023top-5%

Reliable application of machine learning-based decision systems in the wild is one of the major challenges currently investigated by the field. A large portion of established approaches aims to detect erroneous predictions by means of assigning confidence scores. This confidence may be obtained by e…

2023

Navigating the Pitfalls of Active Learning Evaluation: A Systematic Framework for Meaningful Performance Assessment

NeurIPS 2023poster

Active Learning (AL) aims to reduce the labeling burden by interactively selecting the most informative samples from a pool of unlabeled data. While there has been extensive research on improving AL query methods in recent years, some studies have questioned the effectiveness of AL compared to emerg…