← Search

Hans Arno Jacobsen

2 accepted papers

2025

Dynamic Loss-Based Sample Reweighting for Improved Large Language Model Pretraining

ICLR 2025poster

Pretraining large language models (LLMs) on vast and heterogeneous datasets is crucial for achieving state-of-the-art performance across diverse downstream tasks. However, current training paradigms treat all samples equally, overlooking the importance or relevance of individual samples throughout t…

Cited by 0SourcePDFScholar
2025

MESS+: Dynamically Learned Inference-Time LLM Routing in Model Zoos with Service Level Guarantees

NeurIPS 2025poster

Open-weight large language model (LLM) zoos provide access to numerous high-quality models, but selecting the appropriate model for specific tasks remains challenging and requires technical expertise. Most users simply want factually correct, safe, and satisfying responses without concerning themsel…

Cited by 0SourceScholar