← Search

Zixuan Ren

3 accepted papers

2026

Enough is as good as a feast: A Comprehensive Analysis of How Reinforcement Learning Mitigates Task Conflicts in LLMs

ICLR 2026poster

Model merging plays a crucial role in consolidating multiple specialized models into a single, unified model, especially in the era of large language models (LLMs). Recent research has primarily focused on developing strategies to enhance merging performance with the trained models, while the impact…

Cited by 0SourceScholar
2026

LLMs are Single-threaded Reasoners: Demystifying the Working Mechanism of Soft Thinking

ICLR 2026poster

Human cognition naturally engages with abstract and fluid concepts, whereas existing reasoning models often rely on generating discrete tokens, potentially constraining their expressive capabilities. Recent advancements aim to address this limitation by enabling large language models (LLMs) to gener…

Cited by 0SourceScholar
2023

Towards Informative Open-ended Text Generation with Dynamic Knowledge Triples

EMNLP 2023long findings

Pretrained language models (PLMs), especially large language models (LLMs) demonstrate impressive capabilities in open-ended text generation. While our statistical results show that LLMs often suffer from over-concentrated information, where the generated texts overly focus on the given prompt and f…

Cited by 0SourceScholar