← Search

Junjie Yao

1 accepted papers

2025

An Analysis for Reasoning Bias of Language Models with Small Initialization

ICML 2025spotlight

Transformer-based Large Language Models (LLMs) have revolutionized Natural Language Processing by demonstrating exceptional performance across diverse tasks. This study investigates the impact of the parameter initialization scale on the training behavior and task preferences of LLMs. We discover th…

Cited by 1SourcePDFScholar