← Search

Andreas Hotho

4 accepted papers

2025

LLäMmlein: Transparent, Compact and Competitive German-Only Language Models from Scratch

ACL 2025long

We transparently create two German-only decoder models, LLäMmlein 120M and 1B, from scratch and publish them, along with the training data, for the (German) NLP research community to use. The model training involved several key steps, including data preprocessing/filtering, the creation of a German…

2021

Do Different Deep Metric Learning Losses Lead to Similar Learned Features?

ICCV 2021poster

Recent studies have shown that many deep metric learning loss functions perform very similarly under the same experimental conditions. One potential reason for this unexpected result is that all losses let the network focus on similar image regions or properties. In this paper, we investigate this b…

Cited by 13PDFcodeScholar