← Search

Nikos Karampatziakis

7 accepted papers

2024

LoftQ: LoRA-Fine-Tuning-aware Quantization for Large Language Models

ICLR 2024oral

Quantization is an indispensable technique for serving Large Language Models (LLMs) and has recently found its way into LoRA fine-tuning (Dettmers et al., 2023). In this work we focus on the scenario where quantization and LoRA fine- tuning are applied together on a pre-trained model. In such cases…

2017

Gradient Coding: Avoiding Stragglers in Distributed Learning

ICML 2017poster

We propose a novel coding theoretic framework for mitigating stragglers in distributed learning. We show how carefully replicating data blocks and coding across gradients can provide tolerance to failures and stragglers for synchronous Gradient Descent. We implement our schemes in python (using MPI)…

Cited by 590SourcePDFScholar