← Search

Michael Shavlovsky

2 accepted papers

2025

COS-DPO: Conditioned One-Shot Multi-Objective Fine-Tuning Framework

UAI 2025

In LLM alignment and many other ML applications, one often faces the *Multi-Objective Fine-Tuning* (MOFT) problem, *i.e.*, fine-tuning an existing model with datasets labeled w.r.t. different objectives simultaneously. To address the challenge, we propose a *Conditioned One-Shot* fine-tuning framewo

2024

Accelerating Sinkhorn algorithm with sparse Newton iterations

ICLR 2024poster

Computing the optimal transport distance between statistical distributions is a fundamental task in machine learning. One remarkable recent advancement is entropic regularization and the Sinkhorn algorithm, which utilizes only matrix scaling and guarantees an approximated solution with near-linear r…

Cited by 5SourcePDFScholar