2025
Understanding LLM Behaviors via Compression: Data Generation, Knowledge Acquisition and Scaling Laws
NeurIPS 2025spotlight
Large Language Models (LLMs) have demonstrated remarkable capabilities across numerous tasks, yet principled explanations for their underlying mechanisms and several phenomena, such as scaling laws, hallucinations, and related behaviors, remain elusive. In this work, we revisit the classical relatio…