2025
An Analysis for Reasoning Bias of Language Models with Small Initialization
ICML 2025spotlight
Transformer-based Large Language Models (LLMs) have revolutionized Natural Language Processing by demonstrating exceptional performance across diverse tasks. This study investigates the impact of the parameter initialization scale on the training behavior and task preferences of LLMs. We discover th…