MaZO: Masked Zeroth-Order Optimization for Multi-Task Fine-Tuning of Large Language Models
Large language models have demonstrated exceptional capabilities across diverse tasks, but their fine-tuning demands significant memory, posing challenges for resource-constrained environments. Zeroth-order (ZO) optimization provides a memory-efficient alternative by eliminating the need for backpro