2025
From Decoupling to Adaptive Transformation: a Wider Optimization Space for PTQ
ICLR 2025poster
Post-Training low-bit Quantization (PTQ) is useful to accelerate DNNs due to its high efficiency, the current SOTAs of which mostly adopt feature reconstruction with self-distillation finetuning. However, when bitwidth goes to be extremely low, we find the current reconstruction optimization space i…