← Search

Peijian Gu

3 accepted papers

2026

FineRef: Fine-Grained Error Reflection and Correction for Long-Form Generation with Citations

AAAI 2026technical

Generating with citations is crucial for trustworthy Large Language Models (LLMs), yet even advanced LLMs often produce mismatched or irrelevant citations. Existing methods over-optimize citation fidelity while overlooking relevance to the user query, which degrades answer quality and robustness in

Cited by 0SourcePDFScholar
2025

Improve Safety Training of Large Language Models with Safety-Critical Singular Vectors Localization

ACL 2025long

The rapid advancement of large language models (LLMs) has brought about increased concerns regarding their safety, especially as adversaries develop jailbreak techniques to bypass LLMs’ safety mechanism. Although recent work on safety training with modules such as low-rank adaptation (LoRA) to resis…

2023

IAEval: A Comprehensive Evaluation of Instance Attribution on Natural Language Understanding

EMNLP 2023long findings

Instance attribution (IA) aims to identify the training instances leading to the prediction of a test example, helping researchers understand the dataset better and optimize data processing. While many IA methods have been proposed recently, how to evaluate them still remains open. Previous evaluati…

Cited by 0SourceScholar