← Search

Xingjin Wang

2 accepted papers

2025

Learning Dynamics in Continual Pre-Training for Large Language Models

ICML 2025oral

Continual Pre-Training (CPT) has become a popular and effective method to apply strong foundation models to specific downstream tasks. In this work, we explore the **learning dynamics** throughout the CPT process for large language models (LLMs). We specifically focus on how general and downstream…

Cited by 0SourcePDFScholar
2023

LDM$^2$: A Large Decision Model Imitating Human Cognition with Dynamic Memory Enhancement

EMNLP 2023long findings

With the rapid development of large language models (LLMs), it is highly demanded that LLMs can be adopted to make decisions to enable the artificial general intelligence. Most approaches leverage manually crafted examples to prompt the LLMs to imitate the decision process of human. However, design…

Cited by 0SourceScholar