2026
Zero Sum SVD: Balancing Loss Sensitivity for Low Rank LLM Compression
ICML 2026poster
Advances in large language models have driven strong performance across many tasks, but their memory and compute costs still hinder deployment. SVD-based compression reduces storage and can speed up inference via low-rank factors, yet performance depends on how rank is allocated under a global compr…