← Search

Yanzhu Guo

4 accepted papers

2025

Do Large Language Models have an English Accent? Evaluating and Improving the Naturalness of Multilingual LLMs

ACL 2025long

Current Large Language Models (LLMs) are predominantly designed with English as the primary language, and even the few that are multilingual tend to exhibit strong English-centric biases. Much like speakers who might produce awkward expressions when learning a second language, LLMs often generate un…

2024

The Curious Decline of Linguistic Diversity: Training Language Models on Synthetic Text

NAACL 2024findings

This study investigates the consequences of training language models on synthetic data generated by their predecessors, an increasingly prevalent practice given the prominence of powerful generative models. Diverging from the usual emphasis on performance metrics, we focus on the impact of this trai…

2023

Automatic Analysis of Substantiation in Scientific Peer Reviews

EMNLP 2023long findings

With the increasing amount of problematic peer reviews in top AI conferences, the community is urgently in need of automatic quality control measures. In this paper, we restrict our attention to substantiation --- one popular quality aspect indicating whether the claims in a review are sufficiently…

Cited by 0SourcecodeScholar
2022

Questioning the Validity of Summarization Datasets and Improving Their Factual Consistency

EMNLP 2022main

The topic of summarization evaluation has recently attracted a surge of attention due to the rapid development of abstractive summarization systems. However, the formulation of the task is rather ambiguous, neither the linguistic nor the natural language processing communities have succeeded in givi…

Cited by 9SourcePDFScholar