2024
Shears: Unstructured Sparsity with Neural Low-rank Adapter Search
NAACL 2024industry
Recently, several approaches successfully demonstrated that weight-sharing Neural Architecture Search (NAS) can effectively explore a search space of elastic low-rank adapters (LoRA), allowing the parameter-efficient fine-tuning (PEFT) and compression of large language models. In this paper, we intr…