← Search

Lukas Rauch

6 accepted papers

2026

Cleaning the Pool: Progressive Filtering of Unlabeled Pools in Deep Active Learning

CVPR 2026

Existing active learning (AL) strategies capture fundamentally different notions of data value, e.g., uncertainty or representativeness. Consequently, the effectiveness of strategies can vary substantially across datasets, models, and even AL cycles. Committing to a single strategy risks suboptimal

Cited by 0SourceScholar
2026

Hashing-Baseline: Rethinking hashing in the age of pretrained models

ICASSP 2026poster

Information retrieval with compact binary embeddings, also referred to as hashing, is crucial for scalable fast search applications, yet state-of-the-art hashing methods require expensive, scenario-specific training. In this work, we introduce Hashing-Baseline, a strong training-free hashing method…

Cited by 0SourcePDFScholar
2026

Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification

ICLR 2026poster

Although probing frozen models has become a standard evaluation paradigm, self-supervised learning in audio defaults to fine-tuning when pursuing state-of-the-art on AudioSet. A key reason is that global pooling creates an information bottleneck causing linear probes to misrepresent the embedding qu…

Cited by 0SourceScholar
2025

BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics

ICLR 2025spotlight

Deep learning (DL) has greatly advanced audio classification, yet the field is limited by the scarcity of large-scale benchmark datasets that have propelled progress in other domains. While AudioSet is a pivotal step to bridge this gap as a universal-domain dataset, its restricted accessibility and…

2024

dopanim: A Dataset of Doppelganger Animals with Noisy Annotations from Multiple Humans

NeurIPS 2024poster

Human annotators typically provide annotated data for training machine learning models, such as neural networks. Yet, human annotations are subject to noise, impairing generalization performances. Methodological research on approaches counteracting noisy annotations requires corresponding datasets f…