← Search

Ali Abbasi

5 accepted papers

2026

ON THE ROLE OF IMPLICIT REGULARIZATION OF STOCHASTIC GRADIENT DESCENT IN GROUP ROBUSTNESS

ICLR 2026poster

Training with stochastic gradient descent (SGD) at moderately large learning rates has been observed to improve robustness against spurious correlations, strong correlation between non-predictive features and target labels. Yet, the mechanism underlying this effect remains unclear. In this work, we…

Cited by 0SourcecodeScholar
2026

Zero Sum SVD: Balancing Loss Sensitivity for Low Rank LLM Compression

ICML 2026poster

Advances in large language models have driven strong performance across many tasks, but their memory and compute costs still hinder deployment. SVD-based compression reduces storage and can speed up inference via low-rank factors, yet performance depends on how rank is allocated under a global compr…

Cited by 0SourceScholar
2025

MCNC: Manifold-Constrained Reparameterization for Neural Compression

ICLR 2025poster

The outstanding performance of large foundational models across diverse tasks, from computer vision to speech and natural language processing, has significantly increased their demand. However, storing and transmitting these models poses significant challenges due to their massive size (e.g., 750GB…

2024

BrainWash: A Poisoning Attack to Forget in Continual Learning

CVPR 2024poster

Continual learning has gained substantial attention within the deep learning community offering promising solutions to the challenging problem of sequential learning. Yet a largely unexplored facet of this paradigm is its susceptibility to adversarial attacks especially with the aim of inducing forg…

2023

PRANC: Pseudo RAndom Networks for Compacting Deep Models

ICCV 2023poster

We demonstrate that a deep model can be reparametrized as a linear combination of several randomly initialized and frozen deep models in the weight space. During training, we seek local minima that reside within the subspace spanned by these random models (i.e., `basis' networks). Our framework, PRA…

Cited by 12PDFcodeScholar