2024
Compensate Quantization Errors: Make Weights Hierarchical to Compensate Each Other
NAACL 2024findings
Emergent Large Language Models (LLMs) use their extraordinary performance and powerful deduction capacity to discern from traditional language models. However, the expenses of computational resources and storage for these LLMs are stunning, quantization then arises as a trending conversation. To add…