← Search

Sifeng SHANG

1 accepted papers

2026

Fine-tuning Quantized Neural Networks with Zeroth-order Optimization

ICLR 2026poster

As the size of large language models grows exponentially, GPU memory has become a bottleneck for adapting these models to downstream tasks. In this paper, we aim to push the limits of memory-efficient training by minimizing memory usage on model weights, gradients, and optimizer states, within a uni…

Cited by 0SourcecodeScholar