2025
Minimal Ranks, Maximum Confidence: Parameter-efficient Uncertainty Quantification for LoRA
EMNLP 2025
Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of large language models by decomposing weight updates into low-rank matrices, significantly reducing storage and computational overhead. While effective, standard LoRA lacks mechanisms for uncertainty quantification, leading to over